NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Extreme-value forest fire prediction A study of the Loss Function in an Ordinality Scheme
TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning
MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents
Quantitative Evaluation of the Severity of Posttraumatic Stress Disorder through Transfer…
F-GRPO: Don't Let Your Policy Learn the Obvious and Forget the Rare
Counterfactual Explanations for Hypergraph Neural Networks
VEN-VL: A Visual Ensemble MoE Framework for Effective and Efficient Multi-Modal Understan…
Krause Synchronization Transformers
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tun…
Learning in Low-Dimensional Subspaces: Orthogonal Bottlenecks for Reinforcement Learning
Red-Teaming Claude Opus and ChatGPT-based Security Advisors for Trusted Execution Environ…
Pixelwise Uncertainty Quantification of Accelerated MRI Reconstruction
Future-KL Regularized GRPO: Process-Level Credit Assignment from $f$-Divergence Regulariz…
CARL-CXR: Continual Adapter-Based Routing for Task-Unknown Chest Radiograph Classification
QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis
MoBiQuant: Mixture-of-Bits Quantization for Token-Adaptive Any-Precision LLM
Everything at Every Scale: Scale-Invariant Diffusion with Continuous Super-Resolution
AI-Assisted Systematization for Evaluating GenAI Systems
Non-Invasive Reconstruction of Intracranial EEG Across the Deep Temporal Lobe from Scalp …
Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Rein…
RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with…
MultiPUFFIN: A Multimodal Domain-Constrained Foundation Model for Molecular Property Pred…
Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation
Learning from Trials and Errors: Reflective Test-Time Planning for Embodied LLMs
PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Rein…
UtilityMax Prompting: A Formal Framework for Multi-Objective Large Language Model Tasks
Language Models Need Sleep
PowerFlow: Unlocking the Dual Nature of LLMs via Principled Distribution Matching