Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Evidence-Aligned Entity Verification for Hallucination Detection in Retrieval-Augmented G…
arXiv cs.AI Research & Papers
From Concentration to Differentiation and Back: Routing Effective Rank in MoE Reasoning C…
arXiv cs.AI Research & Papers
SchemeArena: Factorized Stress Testing of Scheming in LLM Agents
arXiv cs.AI Research & Papers
From Simulated Citizens to Simulated Deliberation: Challenges in Representation and Inter…
arXiv cs.AI Research & Papers
AutoKD: Autonomous Knowledge Discovery
arXiv cs.AI Research & Papers
Modus Tollens and Counterfactuals and Counterfactual Reasoning Based on Three Types of Ne…
arXiv cs.AI Research & Papers
SAP: State-Guided Data Synthesis with Argument Provenance for Multi-Turn Tool Use
arXiv cs.AI Research & Papers
Building Trustworthy Graph-Agentic RAG for Social Good: Architectures, Failure Propagatio…
arXiv cs.AI Research & Papers
Simulating the Marginal Green Contribution of AI Modules in a Smart-Agriculture Platform:…
arXiv cs.AI Research & Papers
A Data-Driven Framework for Identifying and Prioritizing RPA Opportunities in Healthcare …
arXiv cs.AI Research & Papers
Diffusion models for eye-gaze trajectory generation using position and velocity represent…
arXiv cs.AI Research & Papers
Predicting Wind Turbine Power Using Machine Learning and Weather Forecasts
arXiv cs.AI Research & Papers
Towards Embodied Air-Ground Cooperative Object Search: Benchmark, Dataset and Agentic Met…
arXiv cs.AI Research & Papers
DGCPath: Distribution-Aware Generative Contrastive Framework for Self-supervised Path Rep…
arXiv cs.AI Research & Papers
SE-GoS: Self-Evolving Graph-of-Skills for Skill Library at Scale
arXiv cs.AI Research & Papers
Qiushi Engine on AstaBench E2E-Bench-Hard
arXiv cs.AI Research & Papers
AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems
arXiv cs.AI Research & Papers
Improving Proficiency and Efficiency of Android GUI Agents via Self-Generating Tool Actio…
arXiv cs.AI Research & Papers
Risk Is Not Review Value: Wrong-Answer Exposure Under Bounded Review Budgets
arXiv cs.AI Research & Papers
SCIRIGOR:Evaluating Open-Ended Scientific Analysis Beyond Final Scores
arXiv cs.AI Research & Papers
Key Path Identification for Resolving Knowledge Conflicts via SAE-based Steering
arXiv cs.AI Security & Safety
Revoked but Still Authoritative: An Empirical Study of Revocation Enforcement in Agent-Me…
arXiv cs.AI Research & Papers
CLAMP: Constrained Decoding for Vision-Language Embodied Planning
arXiv cs.AI Research & Papers
Data-driven rational function neural networks: a new method for generating analytical mod…
arXiv cs.AI Research & Papers
DART: A DAG-Based Reputation and Incentive Framework via Blockchain-Enabled Governance fo…
arXiv cs.AI Research & Papers
A Tool-Augmented, GPT-4 Chatbot for Real-Time Repository Data Analysis
arXiv cs.AI Research & Papers
AMA: Adaptive Memory via Multi-Agent Collaboration
arXiv cs.AI Research & Papers
Scratchy: Visual-Scratchpad Multimodal Reasoning for Cryptographic Proof Generation in Ea…
arXiv cs.AI Research & Papers
OntologyBench: Can Dense Retrieval Satisfy Structured Biomedical Constraints?
arXiv cs.AI Research & Papers
Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collaps…