Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search…
Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete…
Differentiable Belief-based Opponent Shaping
BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inferen…
ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive In…
BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation
Topological Order in Neural Wavefunctions
Mechanistic origins of catastrophic forgetting: why RL preserves circuits better than SFT?
DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Age…
PrismFlow: Residual Dynamics for Flow Matching in Time-Series Generation
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Secu…
SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow
Locally Coherent, Globally Incoherent: Bounding Compositional Incoherence in Multi-Compon…
PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster…
Hallucination Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Sema…
Who can we trust? LLM-as-a-jury for Comparative Assessment
Continuity and Ordinality Matter: Constraining Time Series Tokens for Effective Time Seri…
TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evalu…
Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories
LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis
It`s All About Speed: AI`s Impact on Workflow in Music Production
Meta-Programming for Linear-time Temporal Answer Set Programming
Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent
Practitioner Beliefs and Behaviors in AI-Enhanced Education: DOT Framework Survey Evidence
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
Ultra-Reduced-Impact-Encased-Logging (URIEL): propose a new method for selective sustaina…
Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation
From GPS Points to Travel Patterns: Flexible and Semantic Trajectory Generation with LLMs