Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Useful Memories Become Faulty When Continuously Updated by LLMs
arXiv cs.AI Research & Papers
Safe to Resume? Breaking Execution Continuity of Agent Execution via Rollback
arXiv cs.AI Research & Papers
What Does an Evaluation License? A Commit-Bound Census of Claim Replay in Inspect Evals
arXiv cs.AI Research & Papers
AGRICAM: A Track-Mounted Crop Pollination Monitoring Robot
arXiv cs.AI Research & Papers
COMPASS: Grounding Composition-Intent Guidance in Unified Multimodal Models
arXiv cs.AI Research & Papers
M2K: Making the Model-Kernel Interface Explicit for Reliable CUDA Kernel Verification
arXiv cs.AI Research & Papers
FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effecti…
arXiv cs.AI Research & Papers
Learning Composable Chains-of-Thought
arXiv cs.AI Research & Papers
Retrieval-Augmented LLM Agents: Learning to Learn from Experience
arXiv cs.AI Research & Papers
WebXSkill: Skill Learning for Autonomous Web Agents
arXiv cs.AI Research & Papers
A Multi-Agent Human-LLM Collaborative Framework for Closed-Loop Scientific Literature Sum…
arXiv cs.AI Research & Papers
PeopleSearchBench: Evaluating AI-Powered People Search Platforms with Criteria-Grounded V…
arXiv cs.AI Research & Papers
Uncertainty Makes It Stable: Curiosity-Driven Quantized Mixture-of-Experts
arXiv cs.AI Research & Papers
Correctness Forensics for Batch Speculative Decoding: Diagnosing the Ragged Tensor Problem
arXiv cs.AI Research & Papers
Disappearing Ink: Obfuscation Breaks N-gram Code Watermarks in Theory and Practice
arXiv cs.AI Research & Papers
MobileDreamer: Generative Sketch World Model for GUI Agent
arXiv cs.AI Research & Papers
Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models
arXiv cs.AI Research & Papers
Proof2Hybrid: Automatic Mathematical Benchmark Synthesis for Proof-Centric Problems
arXiv cs.AI Research & Papers
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
arXiv cs.AI Research & Papers
GUI-PRA: Process Reward Agent for GUI Tasks
arXiv cs.AI Research & Papers
ReGraP-LLaVA: Reasoning enabled Graph-based Personalized Large Language and Vision Assist…
arXiv cs.AI Research & Papers
Personas Differ from Native-Language Generation: Language Pathways Shape LLM Interpersona…
arXiv cs.AI Research & Papers
Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper
arXiv cs.AI Research & Papers
A Composition-Aware Pretraining Framework for Geospatial Foundation Models
arXiv cs.AI Research & Papers
SimGuide: Typed Multi-Context User Representations for Preference-Conditioned Agent Plann…
arXiv cs.AI Research & Papers
Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space
arXiv cs.AI Research & Papers
Fully Distributed GNE Algorithms for Multi-Robot Placement without Consensus on Multiplie…
arXiv cs.AI Research & Papers
Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems
arXiv cs.AI Research & Papers
On the Prospects of Dynamic LLM Conversations in Software Development
arXiv cs.AI Research & Papers
A Mental Model Based Framework of Trust