Explorar

Noticias de IA

22115 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Interv…
arXiv cs.AI Research & Papers
Representation Robustness Under Executable Reasoning Constraints in Large Language Models…
arXiv cs.AI Research & Papers
CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits
arXiv cs.AI Research & Papers
SiGMA: Sign-Guided Merging and Adaptation for Multimodal Continual Instruction Tuning
arXiv cs.AI Research & Papers
Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain
arXiv cs.AI Research & Papers
Autonomous disproofs of the sum-product conjecture over $\mathbb R$ with GPT-5.5 Pro
arXiv cs.AI Research & Papers
JAXBench: Benchmarking Autonomous TPU Kernel Optimization
arXiv cs.AI Research & Papers
Reliability-Aware LLM Alignment from Inconsistent Human Feedback
arXiv cs.AI Research & Papers
Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Serie…
arXiv cs.AI Research & Papers
Mean-to-Score Discrete Diffusion: Posterior-Mean Denoisers for Score Entropy
arXiv cs.AI Research & Papers
VoLN: Vision-Only Long-Horizon Navigation---Paradigm, Benchmark, and Method
arXiv cs.AI Research & Papers
Constrained latent state modeling: A unifying perspective on representation learning unde…
arXiv cs.AI Research & Papers
Cycle-Consistent and Uncertainty-Aware Neural Surrogates for Tokamak Edge Plasmas
arXiv cs.AI Research & Papers
Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervis…
arXiv cs.AI Research & Papers
Probabilistic Residual Learning for Online Recommendations
arXiv cs.AI Research & Papers
Thinkink: 2D Spatial Ink-native Interaction with LLMs
arXiv cs.AI Research & Papers
LeanFlow: A Case Study in Workflow-Driven Lean Autoformalization
arXiv cs.AI Research & Papers
MKEvolve: A Modular Multi-Agent Framework for Kernel Code Generation
arXiv cs.AI Research & Papers
FlowEdit: Information-Theoretic Control of LLM Reasoning Flows for Ill-posed Problems Inv…
arXiv cs.AI Research & Papers
ExecuGraph: A Multi-Agent, Execution-Grounded Framework for Reliable Backend Code Synthes…
arXiv cs.AI Research & Papers
AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge …
arXiv cs.AI Research & Papers
From Errors to Rules: Iterative Prompt Optimization for Text Classification
arXiv cs.AI Research & Papers
Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections
arXiv cs.AI Research & Papers
Isolating LLM Alignment from Regex: Zero Coverage and Metric-Dependent Divergence Under A…
arXiv cs.AI Research & Papers
Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure,…
arXiv cs.AI Research & Papers
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing
arXiv cs.AI Research & Papers
Multimodal Learning for Arcing Detection in Pantograph-Catenary Systems
arXiv cs.AI Research & Papers
Visual Contrastive Self-Distillation
arXiv cs.AI Research & Papers
Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-su…
arXiv cs.AI Research & Papers
OPOD: On-Policy Omni Distillation