Explorar

Noticias de IA

30660 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Agentic coding without the cloud: evaluating open-weight large language models on longitu…
arXiv cs.AI Research & Papers
Playing Along: Learning a Double-Agent Defender for Belief Steering via Theory of Mind
arXiv cs.AI Research & Papers
AREX: Towards a Recursively Self-Improving Agent for Deep Research
arXiv cs.AI Research & Papers
Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry
arXiv cs.AI Research & Papers
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle a…
arXiv cs.AI Research & Papers
Spatially Grounded Concept Bottleneck Models for Trustworthy Breast Ultrasound Diagnosis
arXiv cs.AI Research & Papers
OpenForgeRL: Train Harness-native Agents in Any Environment
arXiv cs.AI Research & Papers
From Agent Failures to Text Policies: What Works and What Breaks
arXiv cs.AI Research & Papers
Multi-Task Learning for Heterogeneous Prediction from Video Game State with Transfer Lear…
arXiv cs.AI Research & Papers
Benchmarking Unlearning for Vision Transformers
arXiv cs.AI Research & Papers
Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for D…
arXiv cs.AI Research & Papers
From Scalars to Time Series: Rethinking Implicit Neural Representations for Time-Varying …
arXiv cs.AI Research & Papers
Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment
arXiv cs.AI Research & Papers
Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts
arXiv cs.AI Research & Papers
MSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamic…
arXiv cs.AI Research & Papers
Delivery, Not Storage: Cue-Anchored Working Memory as a Harness Property for Coding Agents
arXiv cs.AI Research & Papers
SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integr…
arXiv cs.AI Research & Papers
DFAH-Bench: Benchmarking Observable Agent Instability in Financial Decision-Making
arXiv cs.AI Research & Papers
Multimodal Pretraining for Generalizable EEG Representation Learning
arXiv cs.AI Research & Papers
RE-AD: Real-Time Requirement Adherence for Data Labeling
arXiv cs.AI Research & Papers
OPOD: On-Policy Omni Distillation
arXiv cs.AI Research & Papers
Instruct-FD: Can Your Full-Duplex Speech System Follow Turn-Taking Instructions?
arXiv cs.AI Research & Papers
Constrained latent state modeling: A unifying perspective on representation learning unde…
arXiv cs.AI Research & Papers
Beyond Independent Optimization: Compression, MoE Routing, and Quantization Interactions …
arXiv cs.AI Research & Papers
Euclid-MCP: A Model Context Protocol Server for Deterministic Logical Reasoning via Prolog
arXiv cs.AI Research & Papers
Answer-then-Edit: Reasoning Skeleton Editing for Anti-Distillation with Preserved Utility
arXiv cs.AI Research & Papers
Beyond SBDD: Geometric Deep Learning in Polypharmacology and Multi-target Drug Design
arXiv cs.AI Research & Papers
Probabilistic Residual Learning for Online Recommendations
arXiv cs.AI Research & Papers
Refusal-Gated Decoding: Preserving Refusal Behavior Under High-Temperature Sampling
arXiv cs.AI Research & Papers
SPORD: A Simulation-Propose-then-OR-Dispose Approach for Supply Chain Planning