Explorar

Noticias de IA

21863 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
RAID: Towards Robust AI-Generated Image Detection with Bit-Reversed Images
arXiv cs.AI Research & Papers
WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization
arXiv cs.AI Research & Papers
Semantics of Subterfuge: Benchmarking Legal Deception Detection Against General-domain St…
arXiv cs.AI Research & Papers
Evaluating Federated Pre-Training: On the Reliability of Downstream Fine-Tuning and Intri…
arXiv cs.AI Research & Papers
LAWFUL: Law-Aligned Witness for Faithful Use of Latents
arXiv cs.AI Research & Papers
PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learnin…
arXiv cs.AI Research & Papers
TerraNova: A Foundation Model for the Anthropocene
arXiv cs.AI Research & Papers
LLM Framework for Discovering Major Mathematical Conjectures: AI's Quest for the Next Rie…
arXiv cs.AI Research & Papers
ThinkReset: Learnable Intermediate Interface Construction for Bounded-Context Long-Horizo…
arXiv cs.AI Research & Papers
OpenClaw and Ollama in Agentic AI: Toward Fully Autonomous and Scalable AI Agent Systems
arXiv cs.AI Research & Papers
Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Sys…
arXiv cs.AI Research & Papers
Empowering Cross-Domain Sequential Recommendation with Hybrid Tokenization and Serial-Par…
arXiv cs.AI Research & Papers
Scaling Scientific Discovery Environments for Turn-Level Agentic RL
arXiv cs.AI Research & Papers
On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness
arXiv cs.AI Research & Papers
Gated Q-learning: Add Off-Policy Bias to Taste
arXiv cs.AI Research & Papers
ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding
arXiv cs.AI Research & Papers
Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures
arXiv cs.AI Research & Papers
Identifying Informative Environments for Cognition Parameter Inference via Bayesian Exper…
arXiv cs.AI Research & Papers
SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Ac…
arXiv cs.AI Research & Papers
EarlyDx: An Admission-Anchored Benchmark for Open-Ended Generation of Evidence-Supported …
arXiv cs.AI Research & Papers
FairFund-Bench: Evaluating Distributive Bias in LLM Resource Allocation
arXiv cs.AI Research & Papers
ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation
arXiv cs.AI Research & Papers
Creative Integration: A Decidable Criterion of Creativity
arXiv cs.AI Research & Papers
Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Rememb…
arXiv cs.AI Research & Papers
Beyond Component Testing: Validating Agentic AI Systems
arXiv cs.AI Research & Papers
Beyond Retrieval: Analytic Memory for Multimodal Agents
arXiv cs.AI Research & Papers
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
arXiv cs.AI Research & Papers
Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL
arXiv cs.AI Research & Papers
Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
arXiv cs.AI Research & Papers
COntExt: Towards Context-Aware Ontology Extension from Operational Metrics