Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Identifying and Understanding Human Values in Text: A Tailorable LLM-based Architecture
arXiv cs.AI Research & Papers
Sense Representations Are Inducible Interfaces
arXiv cs.AI Research & Papers
EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization …
arXiv cs.AI Research & Papers
On the Origin of Synthetic Information by Means of Steganographic Inheritance
arXiv cs.AI Research & Papers
DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM…
arXiv cs.AI Research & Papers
Discovery Agents for Real-Time Analytics: Toward Proactive Insight Systems
arXiv cs.AI Developer Tooling
Agyn: An Open-Source Platform for AI Agents with Scalable On-Demand Execution, Agent Defi…
arXiv cs.AI Models & Releases
Laguna M.1/XS.2 Technical Report
arXiv cs.AI Research & Papers
ReflexGrad: Within-Episode Failure Recovery in LLM Agents via Progress-Gated Dual-Process…
arXiv cs.AI Research & Papers
Reasoning and Planning with Dynamically Changing Norms
arXiv cs.AI Research & Papers
Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Syst…
arXiv cs.AI Research & Papers
DiagramBank: A Quality-Audited Dataset of Scientific Schematic Diagrams with Multi-Level …
arXiv cs.AI Research & Papers
Do Clinical Models Change Treatment Decisions?
arXiv cs.AI Research & Papers
DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Esc…
arXiv cs.AI Research & Papers
Probability-Entropy Calibration: An Elastic Indicator for Adaptive Fine-tuning
arXiv cs.AI Research & Papers
Path Channels and Plan Extension Kernels: a Mechanistic Description of Planning in a Soko…
arXiv cs.AI Research & Papers
STFlow: Data-Coupled Flow Matching for Geometric Trajectory Simulation
arXiv cs.AI Research & Papers
PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience …
arXiv cs.AI Research & Papers
Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resoluti…
arXiv cs.AI Research & Papers
A Query Engine for the Agents
arXiv cs.AI Research & Papers
GradientStabilizer:Fix the Norm, Not the Gradient
arXiv cs.AI Research & Papers
Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM A…
arXiv cs.AI Research & Papers
A Fixed-Budget, Cluster-Aware Standard for LLM-as-a-Judge Evaluation: A Multi-Hop RAG Str…
arXiv cs.AI Research & Papers
EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents
arXiv cs.AI Research & Papers
TCP-MCP: Landscape-Guided Co-Evolution of Prompts and Communication Topologies for Multi-…
arXiv cs.AI Research & Papers
Text2Model: Modeling Copilots for Text-to-Model Translation
arXiv cs.AI Research & Papers
DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision…
arXiv cs.AI Research & Papers
SuiChat-CN: Benchmarking Contextual Suicide Risk Assessment in Chinese Group Chats
arXiv cs.AI Research & Papers
STAB: Specification-driven Testing for Algorithmic Bottlenecks
arXiv cs.AI Research & Papers
An Empirical Audit of k-NAF Budget Accounting for Anchored Decoding