Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries Duri…
arXiv cs.AI Research & Papers
Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language Agents
arXiv cs.AI Research & Papers
Business Arena: Benchmarking LLM Agents in a Realistic Marketplace
arXiv cs.AI Research & Papers
Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasonin…
arXiv cs.AI Research & Papers
FinTrace: Holistic Trajectory-Level Evaluation of LLM Tool Calling for Long-Horizon Finan…
arXiv cs.AI Research & Papers
Deferred Audio Pruning with Local Audio-Visual Dynamics for Omni-LLMs
arXiv cs.AI Research & Papers
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
arXiv cs.AI Research & Papers
ToolVision: Learning When and How to Use Visual Tools with Capability-Aligned Supervision
arXiv cs.AI Research & Papers
Metanormative Theory for RL-Based Moral Agents
arXiv cs.AI Research & Papers
Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump…
arXiv cs.AI Research & Papers
What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files
arXiv cs.AI Research & Papers
DRBENCHER: Can Your Agent Identify the Entity, Retrieve Its Properties and Do the Math?
arXiv cs.AI Research & Papers
CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions a…
arXiv cs.AI Research & Papers
Time-Series Forecasting in Safety-Critical Environments: An Open-Source Package for EU-AI…
arXiv cs.AI Research & Papers
Mitigating Over-Personalization in LLMs via Structured Memory
arXiv cs.AI Research & Papers
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
arXiv cs.AI Research & Papers
LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving
arXiv cs.AI Research & Papers
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
arXiv cs.AI Research & Papers
Branch2Skill: Efficient Skill Evolution Through Reasoning Trees
arXiv cs.AI Research & Papers
From Operational Design Domain to Action: A Systematic Behavioral Taxonomy for Autonomous…
arXiv cs.AI Research & Papers
StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoni…
arXiv cs.AI Research & Papers
Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective
arXiv cs.AI Research & Papers
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalig…
arXiv cs.AI Research & Papers
A Statistical Framework for Auditing Behavioral Dependence and Induced Bias in LLM Judges
arXiv cs.AI Research & Papers
Culturally Situated AI Safety for Youth: Saudi Arabian Perspectives of Youth, Parents and…
arXiv cs.AI Research & Papers
DualCert: A Solver for the Traveling Salesman Problem with Constraint-Coupled Learning
arXiv cs.AI Research & Papers
Continual Visual Anomaly Detection on the Edge: Benchmark and Efficient Solutions
arXiv cs.AI Research & Papers
Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language …
arXiv cs.AI Security & Safety
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajector…
arXiv cs.AI Research & Papers
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver