Explaining Weather Bulletins via ILP
Explorar
Noticias de IA
30666 elementos — filtrados, clasificados y sin duplicados
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
Evaluating Risks in Weak-to-Strong Alignment: A Bias-Variance Perspective
Diagnosing Pathological Chain-of-Thought in Reasoning Models
Response drift across frontier large language models
StackingNet: Collective Inference Across Independent AI Foundation Models
RE-AD: Real-Time Requirement Adherence for Data Labeling
WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance
A Counterfactual Cause in Situation Calculus
Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models
Confidently Deceptive: How Confidence Amplifies the Risk of LLM Deception
From Resource Flow to Executable Tests: Petri-Net-Guided LLM Test Generation for Concurre…
Visual Contrastive Self-Distillation
Barzilai-Borwein Fails Superlinear Convergence on an Open Set of Quadratics for Every Dim…
Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure,…
Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections
The Storyteller in the Model: Narrative Pattern Inheritance, Escalation Dynamics, and Ali…
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing
Is MoE Routing a Huffman Code? Discovering the Frequency-Diversity Law in Chain-of-Thought
LLM-INSTRUCT at UZH Shared Task 2026: Constraint-Aware Retrieval and Selective Debate for…
HyWorldVLA: A Vision-Language-Action Model with Hybrid World Modeling for Autonomous Driv…
Moir: Let the Model Direct Its Own Story for Robust Cross-Domain Knowledge Editing
Knowledge Injection Exists in MoE? Exploring Expert-Aware Contrast Decoding in MoE for Mi…
Break Through the Compression Bottleneck: From Theory to Practice
Cycle-Consistent and Uncertainty-Aware Neural Surrogates for Tokamak Edge Plasmas
Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Gh…
Thinkink: 2D Spatial Ink-native Interaction with LLMs
OpenForgeRL: Train Harness-native Agents in Any Environment