DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training
Explorar
Noticias de IA
21813 elementos — filtrados, clasificados y sin duplicados
HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs
Post-Training LLMs as Better Decision-Making Agents: A Regret-Minimization Approach
OLG++: A Semantic Extension of Obligation Logic Graph
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
Unifying and Optimizing Data Values for Selection via Sequential Decision-Making
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Vid…
Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length …
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
On Efficient Scaling of GNNs via IO-Aware Layers Implementations
Fine-grained Verification via Diagnostic Reasoning Supervision for Aspect Sentiment Tripl…
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Info…
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation w…
The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of La…
Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Mod…
Scaling Higher-Order Graph Learning with Maximal Clique Complexes
dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Deve…
Appropriateness of Empathy in AI: A Signal-Cost Perspective
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decisio…
SAM for Robust Mitochondria Instance Segmentation in Fluorescence Microscopy
Practical Cross-Band Channel Prediction for AI-RAN via Physics-Guided Deep Unfolding
Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Un…
Correcting Split Selection in Online Decision Trees via Anytime-Valid Inference
What changes after deployment? A survey on On-device Learning in TinyML
Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in …
Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration
MindVoice: Reconstructing Intelligible Speech from Non-invasive Neural Signals with Pretr…
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Over…
Trust-Region Behavior Blending for On-Policy Distillation