EEG-FM-Bench: A Comprehensive Benchmark for the Systematic Evaluation and Diagnostic Anal…
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Frame-Conditioned Moral Computation in LLaMA 3.1-8B-Instruct: A Mechanistic Interpretabil…
ToolMenuBench: Benchmarking Tool-Menu Filtering Strategies for Reliable and Efficient LLM…
When RAG Hurts: Diagnosing and Mitigating Attention Distraction in Retrieval-Augmented LV…
EffGen: Enabling Small Language Models as Capable Autonomous Agents
SLUM-i: Semi-supervised Learning for Urban Mapping of Informal Settlements and Data Quali…
Learning to Share: Selective Memory for Efficient Parallel Agentic Systems
The algebra of Krom logic programs
A biological vision inspired framework for machine perception of abutting grating illusor…
UniT: Unified Multimodal Chain-of-Thought Test-time Scaling
Revisiting Chebyshev Polynomial and Anisotropic RBF Models for Tabular Regression
QoS-Aware Token Scheduling and Private Data Valuation for Multi-Modal Agentic Networks
WorkflowPerturb: Calibrated Stress Tests for Evaluating Multi-Agent Workflow Metrics
Cross-modal Identity Mapping: Minimizing Information Loss in Modality Conversion via Rein…
An Attention Mechanism for Robust Multimodal Integration in a Global Workspace Architectu…
JADE: Expert-Grounded Dynamic Evaluation for Open-Ended Professional Tasks
RaBiT: Residual-Aware Binarization Training for Accurate and Efficient LLMs
WavSLM: Single-Stream Speech Language Modeling via WavLM Distillation
MAND: Modality-Aware Novelty Detection for Open-World Egocentric Activity Recognition
Rel-Zero: Harnessing Patch-Pair Invariance for Robust Zero-Watermarking Against AI Editing
Parallelizing Tool Execution and LLM Generation for Low-Latency Agent Serving
Edit Knowledge, Not Just Facts via Multi-Step Reasoning over Background Stories
Closing the Auto-Research Loop: An AI Co-Scientist for Production Search Ranking
Predicting model behavior before release by simulating deployment
Investigation by The Atlantic reveals many millions of songs used for AI music training
FusionRS: A Large-Scale RGB-Infrared Remote Sensing Dataset for Dual-Modal Vision-Languag…
Can Europe train a frontier AI model on the compute it owns?
A satellite just learned to find things on its own — here’s what that means
Import AI 461: "Alignment is not on track"; FrontierCode; and synthetic research interns
How Post-Training Shapes Biological Reasoning Models