How Chain-of-Thought Works? Tracing Information Flow from Decoding, Projection, and Activ…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning
Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting …
"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models
Timestep-Aware SVDQuant-GPTQ for W4A4 Quantization of Wan2.2-I2V
Lessons from Penetration Tests on Large-Scale Agent Systems
Two Speeds of Learning: A Representation-Readout Decomposition of Grokking and Double Des…
MiRD: Reliable Set-Valued Prediction for Open-Ended Question Answering via Miscoverage Ri…
Beyond the Data Mesh Illusion: Designing Modern AI-augmented Lakehouses to Bridge the Gap…
Semantic Robustness Probing via Inpainting: An Interactive Tool for Safety-Critical Objec…
Grounding Text Embeddings in Stakeholder Associations
Towards Error-Free EHRs: Reasoning-Intensive Consistency Verification Between Clinical No…
From Feasible to Practical: Pareto-Optimal Synthesis Planning
GraphMind: From Operational Traces to Self-Evolving Workflow Automation
Declarative Data Services: Structured Agentic Discovery for Composing Data Systems
Search-E1: Self-Distillation Drives Self-Evolution in Search-Augmented Reasoning
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
"Give Me BF16 or Give Me Death"? Accuracy-Performance Trade-Offs in LLM Quantization
Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinfo…
XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs
Chain Of Thought Compression: A Theoretical Analysis
UCPO: Uncertainty-Aware Policy Optimization
Implementation of Big Data Analytics for Diabetes Management: Needs Assessment in the Rwa…
Diffuse to Detect: Generative Diffusion Models for Unsupervised IC Anomaly Detection
ReCA: Multi-Shot Long Video Extrapolation via Recursive Context Allocation
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial
Can LLMs Introspect? A Reality Check
Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interac…
Adaptive Multi-prompt Contrastive Network for Few-shot Out-of-distribution Detection