Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration
arXiv cs.AI Research & Papers
Agentic Harnesses: LLM-Driven Verification Layers for Robot Autonomy
arXiv cs.AI Research & Papers
SR-OPSD: Self-Referenced On-Policy Self-Distillation
arXiv cs.AI Research & Papers
REIN: Bridging the Gap between Reasoning and Reliability via Reflection and Abstention Al…
arXiv cs.AI Research & Papers
Physical AI Governance: From Theory to Practice Across Life Cycle
arXiv cs.AI Research & Papers
Bridging the Evaluation Gap: Standardized Benchmarks for Multi-Objective Search
arXiv cs.AI Research & Papers
AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Aut…
arXiv cs.AI Research & Papers
HyperFake: Hyperspectral Reconstruction and Attention-Guided Analysis for Advanced Deepfa…
arXiv cs.AI Research & Papers
EgoBrain: Synergizing Minds and Eyes For Human Action Understanding
arXiv cs.AI Research & Papers
The Cost of Language: Centroid Erasure Exposes and Exploits Modal Competition in Multimod…
arXiv cs.AI Research & Papers
Decided Upstream, Written Late: Locating and Pricing the Cross-Lingual Refusal Circuit of…
arXiv cs.AI Research & Papers
Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?
arXiv cs.AI Research & Papers
Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Op…
arXiv cs.AI Research & Papers
KGCache: Amortized Subgraph Retrieval for KG Reasoning with LLMs
arXiv cs.AI Research & Papers
Critic-Free Deep Reinforcement Learning for Maritime Coverage Path Planning on Irregular …
arXiv cs.AI Research & Papers
CyberAGENTS: Structured Autonomy for Agentic Gamified Learning in Cybersecurity
arXiv cs.AI Research & Papers
Contamination Means Overestimation? A Fine-Grained Empirical Study in Code Intelligence
arXiv cs.AI Research & Papers
ATLASFusion: Aggregation Tracking with Location-Aware Sparse Fusion for Robust Spatio-Tem…
arXiv cs.AI Research & Papers
Rethinking KV Cache Eviction via a Unified Information-Theoretic Objective
arXiv cs.AI Research & Papers
In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
arXiv cs.AI Research & Papers
Your Prompt Is Not the Only Prompt: How Much Do LLMs Weight Structured-Output Schema Desc…
arXiv cs.AI Research & Papers
Mitigating Over-Personalization in LLMs via Structured Memory
arXiv cs.AI Research & Papers
Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning
arXiv cs.AI Research & Papers
LLMVisor: A Real-Time Latency Attribution Model for Multi-Tenant LLM Serving
arXiv cs.AI Research & Papers
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
arXiv cs.AI Research & Papers
Estimating Uncertainty in Galaxy Morphology Classification
arXiv cs.AI Research & Papers
What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files
arXiv cs.AI Research & Papers
Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production
arXiv cs.AI Research & Papers
SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests
arXiv cs.AI Research & Papers
Branch2Skill: Efficient Skill Evolution Through Reasoning Trees