Break Through the Compression Bottleneck: From Theory to Practice
Explorar
Noticias de IA
30663 elementos — filtrados, clasificados y sin duplicados
A New Well-Supported Semantics for Description Logic Programs
Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought
Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mi…
Drive As You Like: Multi-Head Diffusion with Reinforcement Learning for Personalized Driv…
LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for…
OpenForgeRL: Train Harness-native Agents in Any Environment
Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Gh…
EmoAgent-R1: Towards Multimodal Emotion Understanding with Reinforcement Learning-based D…
Robust Critics: Defending LLMs Against Multi-Turn Attacks
Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing
Expectation Alignment of Language Models for Real-World User Expectations
RUMBA: Russian User Memory Benchmark
The Boundaries of Automation: A Theory of Persistent Human Participation
LinearARD: Linear-Memory Attention Distillation for RoPE Restoration
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle a…
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry
KeySI: An Interaction Framework for Tuning Text Embeddings Based on Human Feedback
Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective
From Atoms to Entropy: Optimal Noise Allocation for Diffusion Training in the Convex Regi…
StackingNet: Collective Inference Across Independent AI Foundation Models
Diagnosing Pathological Chain-of-Thought in Reasoning Models
Agentic coding without the cloud: evaluating open-weight large language models on longitu…
AREX: Towards a Recursively Self-Improving Agent for Deep Research
M$^3$-Gen: Interpretable Multimodal Generation of Gene Expression Profiles Using Clinical…
Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for D…
Euclid-MCP: A Model Context Protocol Server for Deterministic Logical Reasoning via Prolog
Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment
Multimodal Pretraining for Generalizable EEG Representation Learning