Attention Degradation, Function Token Anchoring, and the Limits of Attention-Based Interv…
Explorar
Noticias de IA
22115 elementos — filtrados, clasificados y sin duplicados
Representation Robustness Under Executable Reasoning Constraints in Large Language Models…
CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits
SiGMA: Sign-Guided Merging and Adaptation for Multimodal Continual Instruction Tuning
Telco-GAIA: Bilingual Benchmark for Agents in Telecom Domain
Autonomous disproofs of the sum-product conjecture over $\mathbb R$ with GPT-5.5 Pro
JAXBench: Benchmarking Autonomous TPU Kernel Optimization
Reliability-Aware LLM Alignment from Inconsistent Human Feedback
Attention-based Experience Replay Framework for Continual Learning of Agnostic Time Serie…
Mean-to-Score Discrete Diffusion: Posterior-Mean Denoisers for Score Entropy
VoLN: Vision-Only Long-Horizon Navigation---Paradigm, Benchmark, and Method
Constrained latent state modeling: A unifying perspective on representation learning unde…
Cycle-Consistent and Uncertainty-Aware Neural Surrogates for Tokamak Edge Plasmas
Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervis…
Probabilistic Residual Learning for Online Recommendations
Thinkink: 2D Spatial Ink-native Interaction with LLMs
LeanFlow: A Case Study in Workflow-Driven Lean Autoformalization
MKEvolve: A Modular Multi-Agent Framework for Kernel Code Generation
FlowEdit: Information-Theoretic Control of LLM Reasoning Flows for Ill-posed Problems Inv…
ExecuGraph: A Multi-Agent, Execution-Grounded Framework for Reliable Backend Code Synthes…
AISE-Bench: A Full-Cycle Curated Benchmark for Information Seeking on Academic Knowledge …
From Errors to Rules: Iterative Prompt Optimization for Text Classification
Compact Latent Coordination for Autonomous Vehicles at Unsignalized Intersections
Isolating LLM Alignment from Regex: Zero Coverage and Metric-Dependent Divergence Under A…
Artificial Epanorthosis: Why large language models overuse a classical rhetorical figure,…
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing
Multimodal Learning for Arcing Detection in Pantograph-Catenary Systems
Visual Contrastive Self-Distillation
Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-su…
OPOD: On-Policy Omni Distillation