Explorar

Noticias de IA

22114 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Reinforcement Learning for Flow-Matching Policies with Density Transport
arXiv cs.AI Research & Papers
Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents
arXiv cs.AI Research & Papers
Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reaso…
arXiv cs.AI Research & Papers
EinSort: Sorting is All We Need for Tensorizing LLM
arXiv cs.AI Research & Papers
When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manip…
arXiv cs.AI Research & Papers
ActProbe: Action-Space Probe for Early Failure Detection of Generative Robot Policies
arXiv cs.AI Research & Papers
STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for …
arXiv cs.AI Research & Papers
PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems
arXiv cs.AI Research & Papers
Adaptive Loss Balancing for Noise-Robust GRPO in Generative Recommendation
arXiv cs.AI Research & Papers
FlashCP: Load-Balanced Communication-Efficient Context Parallelism for LLM Training
arXiv cs.AI Research & Papers
AgentTrust: A Self-Improving Trust Layer for AI-Agent Actions
arXiv cs.AI Research & Papers
PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR
arXiv cs.AI Research & Papers
Sycophancy as a Multilingual Alignment Failure: How Safety Degrades Across Languages, Top…
arXiv cs.AI Research & Papers
Distilling LLM Reasoning into an Interpretable Policy Tree for Human-AI Collaboration
arXiv cs.AI Research & Papers
Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models
arXiv cs.AI Research & Papers
Hacking Generative Perplexity: Why Unconditional Text Evaluation Needs Distributional Met…
arXiv cs.AI Research & Papers
InA-Probe: Instruction-Aware Active Probing for Time Series Forecasting with LLMs
arXiv cs.AI Research & Papers
PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipula…
arXiv cs.AI Research & Papers
TimpaTeks: Automatic In-place Text Sequence Modification via Diffusion Language Model Ste…
arXiv cs.AI Research & Papers
SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration
arXiv cs.AI Research & Papers
Extending Ontologies: From Dense Embeddings to Hybrid Quantum-Fuzzy Systems
arXiv cs.AI Research & Papers
STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control
arXiv cs.AI Research & Papers
Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy
arXiv cs.AI Research & Papers
ConMem: Structured Memory-Guided Adaptation in Training-Free Multi-Agent Systems
arXiv cs.AI Research & Papers
Structure-Conditioned Actor-Critic Branches for Quality-Diversity Reinforcement Learning
arXiv cs.AI Research & Papers
Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects
arXiv cs.AI Research & Papers
Bridging Expert Knowledge and Automated Feature Engineering via Self-Evolution
arXiv cs.AI Research & Papers
Generative Frontier Planning for Adaptive Peer-Referral Recruitment under Covariate-Depen…
arXiv cs.AI Research & Papers
Q-Delta: Beyond Key-Value Associative State Evolution
arXiv cs.AI Research & Papers
Set-Based Transformer for Atmospheric Compensation in Standoff LWIR Hyperspectral Imaging