Parameter Exploration for RLVR via Variational Learning
Explorar
Noticias de IA
29336 elementos — filtrados, clasificados y sin duplicados
DUET: A Diversity-Quality Duet of Distillation Experts for Two-Step Video Generation
A Rigorous Turing Test: a Foundation for Evaluating Artificial General Intelligence
GRASP: Granularity-Aware Region Alignment and Semantic Prototype Learning for Fine-Graine…
UniMoMo: Expert Merging-Based MoE Acceleration for Large Recommendation Models
Coordinated incentives in AI-generated misinformation governance
Hybrid Neural-Classical Correction for Frozen Time Series Foundation Models: A Comprehens…
Designing for Ethical AI: HCI Feature Considerations to Improve Fairness and User Experie…
LegoLM: Structured Weight Sharing for Large Language Models
SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models
iLTM: Integrated Large Tabular Model
DistillCache: KL-Guided Adaptive KV-Cache Eviction for Memory-Efficient LLM Inference
Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Thre…
Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasonin…
PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling
The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hie…
WuYuEval: A Multi-Level Benchmark for Large Language Models in Solid Waste Management
Ground-Truth Neighborhood Regularization for Reinforcement Learning Post-Training of Time…
Coarse-to-Fine Registration of Jawbone CT and Intraoral Scan Data Using GeDi and ICP with…
COMEX: A Composition-Grounded Benchmark and Learning Framework for Explainable Aesthetic …
SuperCoder: Assembly Program Superoptimization with Large Language Models
KGCache: Amortized Subgraph Retrieval for KG Reasoning with LLMs
Rethinking Self-Evolving Agents: Do We Still Need Prescribed Optimization Pipelines?
Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Age…
Layerwise goal-oriented adaptivity for neural ODEs: an optimal control perspective
SHE: Trajectory-driven Safety Harness Evolution for LLM Agents
Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification
One Adapter Pair per Model: A Universal Activation Interface for Language Models
Idea Search: Guiding Tree Search with Ideas to Explore Diverse Scientific Methods
Agentic Auto-Research is Fuzz Testing