C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries Duri…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language Agents
Business Arena: Benchmarking LLM Agents in a Realistic Marketplace
Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasonin…
FinTrace: Holistic Trajectory-Level Evaluation of LLM Tool Calling for Long-Horizon Finan…
Deferred Audio Pruning with Local Audio-Visual Dynamics for Omni-LLMs
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
ToolVision: Learning When and How to Use Visual Tools with Capability-Aligned Supervision
Metanormative Theory for RL-Based Moral Agents
Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump…
What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files
DRBENCHER: Can Your Agent Identify the Entity, Retrieve Its Properties and Do the Math?
CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions a…
Time-Series Forecasting in Safety-Critical Environments: An Open-Source Package for EU-AI…
Mitigating Over-Personalization in LLMs via Structured Memory
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
Branch2Skill: Efficient Skill Evolution Through Reasoning Trees
From Operational Design Domain to Action: A Systematic Behavioral Taxonomy for Autonomous…
StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoni…
Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalig…
A Statistical Framework for Auditing Behavioral Dependence and Induced Bias in LLM Judges
Culturally Situated AI Safety for Youth: Saudi Arabian Perspectives of Youth, Parents and…
DualCert: A Solver for the Traveling Salesman Problem with Constraint-Coupled Learning
Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions
Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language …
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajector…
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver