Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech…
Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operato…
Uncovering Spontaneous Physics Representations in In-Context Learning
Cortex: Compact Behavior Cloning for Quake with Frozen Visual Features
BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL
Evaluating OpenAI's Privacy Filter: Cross-Lingual, Cross-Domain PII Detection Across 42 B…
The production of meaning in the processing of natural language
A neural operator framework for data-driven discovery of stability and receptivity in phy…
TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions
Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling
Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search
Explainable AI for the EU Right to Explanation: A Systematic Review of the Law-XAI Transl…
A Deployment-Friendly Foundational Framework for Efficient Computational Pathology
Quantifying Hallucinations in Language Language Models on Medical Textbooks
Large Language Models provide support for the parallelogram theory of analogy
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks
GraphCliff: Short-Long Range Gating for Modeling Critical Activity Changes Caused by Subt…
HyperFL: Query-Adaptive Representation Learning for Software Fault Localization
SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling
UrbanAgent: A Tool-Augmented Agent for Cross-System Urban Tasks
When Should Graph Attention Be Sparse? Learning a Per-Edge Tsallis Index
From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Z…
IPPRO: Importance-based Pruning with PRojective Offset for Magnitude-indifferent Structur…
Target-Aligned Fusion for Decision-Sequence Learning under Dynamics Shift
Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Expli…
$\pi$-Attention: Online Efficient Sparse Transformers for Long-Context Modeling
Efficient unsupervised domain adaptation via self-supervised vision transformer and syner…
Compound and Parallel Modes of Tropical Convolutional Neural Networks
When Search Teaches Style: Causal Internalization of Tactical Priors in AlphaZero