When Does Adaptive Guidance Help? Belief-Aware Privileged Distillation for Autonomous Dri…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU
One LR Doesn't Fit All: Heavy-Tail Guided Layerwise Learning Rates for LLMs
Eroding Trust in Real Speech: A Large-Scale Study of Human Audio Deepfake Perception
Natural Language Query to Configuration for Retrieval Agents
2-ASP(Q) programs with weak constraints: Complexity and efficient implementation
Gumbel Machine: Counterfactual Student Writing Generation via Gumbel Noise Steering
VitaBench 2.0: Evaluating Personalized and Proactive Agents in Long-Term User Interactions
Position: AI Safety Requires Effective Controllability
Counteraction-Aware Multi-Teacher On-Policy Distillation for General Capability Recovery …
Developing a Totally Unimodular Linear Program for Optimal Conformance Checking: When and…
From Norms to Indicators (N2I-RAG): An Agentic Retrieval-Augmented Generation Framework f…
Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Mu…
Towards Feedback-to-Plan Decisions for Self-Evolving LLM Agents in CUDA Kernel Generation
It's Not the Capability: Harness Sensitivity Is Non-Monotone Across LLM Agent Tiers
The Attribution Blind Spot: Detecting When Language Models Rely on Memory Rather Than Ret…
Tail-Aware HiFloat4: W4A4 Post-Training Quantization for Wan2.2
UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems
MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Mo…
MobileExplorer: Accelerating On-Device Inference for Mobile GUI Agents via Online Explora…
PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Pr…
From Static Context to Calibrated Interactive RL: Mitigating Distribution Shift in Multi-…
Experiments in Agentic AI for Science
Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems
HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML
Rotation-Invariant Spherical Watermarking via Third-Order SO(3) Representation Coupling
L2Rec: Towards Dual-View Understanding of LLMs for Personalized Recommendation
Cordon-MAS: Defending RAG against Knowledge Poisoning via Information-Flow Control
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Exper…
SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation …