Correcting What You Cannot See: Credit Assignment for Perception Distillation in Multimod…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Group-Reflective Self-Distillation for Agentic Reinforcement Learning
AI and Its Impact on Creativity and Diversity: An Empirical Study of LLM-Generated Produc…
CLQT: A Closed-Loop, Cost-Aware, Strategy-Consistent Benchmark for Diagnostic Evaluation …
TeachArena: Are Language Agents Ready for Realistic Teaching Work?
ICU-Bench:Benchmarking Continual Unlearning in Multimodal Large Language Models
Unifying biomedical knowledge in a modern multimodal graph
When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling
GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation
HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering
OntoTKGE: Ontology-Enhanced Temporal Knowledge Graph Extrapolation
Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems
When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing
Mind the Sim2Real Gap in User Simulation for Agentic Tasks
FinToolBench: Evaluating LLM Agents for Real-World Financial Tool Use
Group Selection as a Safeguard Against AI Substitution
CORE: Collaborative Reasoning via Cross Teaching
Bounded Normative Equivalence in Human-AI Cooperation: Group Behaviour, Not Partner Label…
DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys
Panning for Gold: Expanding Domain-Specific Knowledge Graphs with General Knowledge
LADY: Linear Attention for Autonomous Driving Efficiency without Transformers
SIEVE: Selective Integrity Verification and Escalation for Defending LLM Agents against I…
Knowledge Graph Augmented Large Language Models for Disease Prediction
Near-Optimal Reinforcement Learning for Constrained Recurrence Objectives
Efficiency vs. Alignment: Investigating Safety and Fairness Risks in Parameter-Efficient …
AIC-VDS: Attention-Based In-Context Learning for Joint Velocity Control and Data Collecti…
ProbGuard: Proactive Runtime Monitoring for LLM Agent Safety via Probabilistic Prediction
Topology Enhanced MARL for Multi-Agent Cooperative Decision-Making of CAVs
Rethinking Inference-Time Scaling: Efficiency Limits and Linguistic Signals
Bridging Artificial Intelligence and Power Systems Education Using a Hands-On Executable …