Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Beyond Retrieval: Analytic Memory for Multimodal Agents
arXiv cs.AI Research & Papers
Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Rememb…
arXiv cs.AI Research & Papers
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
arXiv cs.AI Research & Papers
COntExt: Towards Context-Aware Ontology Extension from Operational Metrics
arXiv cs.AI Research & Papers
HenTwin: A Multimodal Digital Twin Framework for Longitudinal Biological State Monitoring…
arXiv cs.AI Research & Papers
Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO
arXiv cs.AI Research & Papers
SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Ac…
arXiv cs.AI Research & Papers
EarlyDx: An Admission-Anchored Benchmark for Open-Ended Generation of Evidence-Supported …
arXiv cs.AI Research & Papers
Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures
arXiv cs.AI Research & Papers
Identifying Informative Environments for Cognition Parameter Inference via Bayesian Exper…
arXiv cs.AI Research & Papers
Scaling Scientific Discovery Environments for Turn-Level Agentic RL
arXiv cs.AI Research & Papers
On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness
arXiv cs.AI Research & Papers
Gated Q-learning: Add Off-Policy Bias to Taste
arXiv cs.AI Research & Papers
FairFund-Bench: Evaluating Distributive Bias in LLM Resource Allocation
arXiv cs.AI Research & Papers
OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems
arXiv cs.AI Research & Papers
Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Sys…
arXiv cs.AI Research & Papers
LLM Framework for Discovering Major Mathematical Conjectures: AI's Quest for the Next Rie…
arXiv cs.AI Research & Papers
ThinkReset: Learnable Intermediate Interface Construction for Bounded-Context Long-Horizo…
arXiv cs.AI Research & Papers
Empowering Cross-Domain Sequential Recommendation with Hybrid Tokenization and Serial-Par…
arXiv cs.AI Research & Papers
ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding
arXiv cs.AI Research & Papers
Semantics of Subterfuge: Benchmarking Legal Deception Detection Against General-domain St…
arXiv cs.AI Developer Tooling
metasignal: A Python Package for Comprehensive Metacognitive Analysis and Decision-Making
arXiv cs.AI Research & Papers
Evaluating Federated Pre-Training: On the Reliability of Downstream Fine-Tuning and Intri…
arXiv cs.AI Research & Papers
LAWFUL: Law-Aligned Witness for Faithful Use of Latents
arXiv cs.AI Research & Papers
SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Indus…
arXiv cs.AI Research & Papers
WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization
arXiv cs.AI Research & Papers
Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates
arXiv cs.AI Research & Papers
A robust association between LLM use and scientific productivity: Assessing stopping-time…
arXiv cs.AI Research & Papers
RAID: Towards Robust AI-Generated Image Detection with Bit-Reversed Images
arXiv cs.AI Research & Papers
PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learnin…