XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
SLMs as Multi-Agent Routers: A Progressive SFT and Reinforcement Learning Approach
Neuro-Symbolic Participation Governance for Verifiable AI Agents in Open Digital Twin Eco…
REIMU: Efficient Heterogeneous Hierarchical Reasoning for SSL-Based Speech Deepfake Detec…
Less Is More: Tuning Configurable Systems with Imperfect Fidelity
Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-…
A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense
Interpretability-Guided Soft Pruning of Attention Heads in Vision Transformers
Inference-Time Policy Alignment for Fair Reinforcement Learning
Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of P…
Logographic Character Visual Pretraining via Semantic-based Contrastive Learning
SymNet: A Multi-Task Network for Joint Radio Map Reconstruction and Transmitter Localizat…
Which Modality Decides? Counterfactual Modality Attribution for Multimodal LLMs
Pretrain on Small Synthetic Data, Scale Large for Free: Symmetry-Aware Foundation Model f…
UEmbed: Unified Sparse and Dense Multimodal Embeddings
Bridging Artificial Intelligence and Power Systems Education Using a Hands-On Executable …
Shared Organizational Memory for Enterprise Coding Agents: System Design and Deployment S…
Can LLM Agents Price Competitively? A Dynamic Multi-Attribute Auction Benchmark for Agent…
Perspectives on Tsallis Statistics for Artificial Intelligence
Co-evolution of social reward and punishment under institutional interventions
PATH-Bench: Path-Dependent Evaluation of Lifelong Agents
Fighting Fire with Fire: On the Feasibility of Protecting Exercises Against AI Cheating
What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs
MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents
Cost-Effective Automated Judging of Natural-Language Mathematical Proofs
Corrigible Assistance in One Round: Pragmatic-Pedagogic Best Response
Width, Memory, and Delay: A Resource Accounting for the Limits of Flat Multi-Agent Systems
Exploring More to Solve More: Boosting Diversity in Text Diffusion Models via Entropy-Bas…
Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiab…
Trajectories That Segment Themselves: Agent-Declared Boundaries as a Training Unit