VISTA: Verifier-Informed Student-to-Teacher Adaptation for On-Policy Self-Distillation
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Effectiveness of IoT and Deep Learning for Detection and Severity Assessment of Postelect…
Cross-Spectral Dense Correspondence for Multimodal Spectral Medical Imaging
CareGraph: An Auditable Hybrid AI Framework for Evidence-Grounded Personalized Longitudin…
Optimal Adversarial Testing: Extracting Honest Test Results from Dishonest Test Takers
Talk in Pieces, See in Whole: Disentangled and Hierarchical Representation Learning in La…
Multimodal Collaborative Debate for Zero-Shot Time Series Reasoning
Blog: Survey of Optimizers
Learning a Size-Weight Frontier for Synthetic-Augmented Inference
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situatio…
Steering Multimodal Large Language Models Decoding for Context-Aware Safety
Describe-Then-Act: Proactive Agent Steering via Distilled Language-Action World Models
Video Generative Models as Geometry Learner
On the Maintenance and Co-evolution of Agent Plugins: An Empirical Study of Claude Code P…
Conformal Uncertainty Quantification Guarantees for Neural Operators
An Enclosed Mode Is a Gauge Choice: Topology Relative to Reach in Certified Code World Mo…
Evaluating the Performance of Large Language Models on GAOKAO Benchmark
How Proper Scoring Rules Shape LLM Forecasting
A Probabilistic Interpretation of KV Cache Eviction
Cross-Session Decomposition Attacks: Scaling Risk and Intent-Aligned Retrieval Defense
When Saying No Makes Better Videos: Designing Dual Gatekeeping for Pedagogically Grounded…
Depth-Aware Pothole Detection Using YOLO and RT-DETR at the Edge
Curvature-Aware Radius Shrinkage for Adaptive Nearest Neighbor Classification
SEGRA: A Structured Experience Guided Reasoning Agent for Property Graph Question Answeri…
Efficient Auto-Interpretability of AI Models in Biology
ReToolSQL: Agentic Reinforcement Learning for Robust Text-to-SQL
Twin Worlds: Equivariance-Based Abstention for Evidence-Grounded Reasoning
Dynamic Alignment Compensation for Hallucination Mitigation in Large Vision-Language Mode…
Compared to What? A Human-Anchored Security Benchmark for LLM-Generated Infrastructure-as…
Understanding and Enforcing Weight Disentanglement in Task Arithmetic