TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
ASTRO: Adaptive Spatio-Temporal Reinforcement Optimization for GNN Powered Anomly Detecti…
LLM Agent Based Renewable Energy Forecasting Using Edge and IoT Data A Review of Solar Wi…
Learning to Reason Efficiently with A* Post-Training
Polynomial Context-Truncation Sensitivity in Autoregressive Language Models: Sequential W…
When Search Becomes Memory: Turning Robot Design Trials into Transferable Skills
Clarify, Abstain or Answer? Strategising in Conversation with Belief-Augmented Generation
TinyFormer: Preserving Tiny Objects in YOLO-DETRHybridReal-time Detectors
Fine-Tuning Over Architectural Complexity: Broad-Coverage PII Detection on PIIBench with …
Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading …
Metropolis-Scale Resilient and Trustworthy Traffic Flow Inference Using Multi-Source Data
MDGMIX: Boundary-Aware Subgraph Mixing for Multi-Domain Graph Pre-Training
Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health …
Selective Test-Time Compute Scaling for Click-Through Rate Prediction via Uncertainty-Tri…
Proper Scoring Rules for Agentic Uncertainty Quantification
Uncertainty Decomposition via Cyclical SG-MCMC and Soft-label Learning for Subjective NLP
SEP-Attack: A Simple and Effective Paradigm for Transfer-Based Textual Adversarial Attack
CoRe-Code: Collaborative Reinforcement Learning for Code Generation
PANDO: Efficient Multimodal AI Agents via Online Skill Distillation
Your Embedding Model is SMARTer Than You Think
VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation
Temporal Concept Drift in Legal Judgment Prediction: Neural Baselines Across Three Epochs…
Explainable Retinal Imaging for Prediction of Multi-Organ Dysfunction in Type 2 Diabetes
Test-Time Deep Thinking to Explore Implicit Rules
Simulating Human Memory with Language Models
Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few Prompts
The Concept Allocation Zone: Tracking How Concepts Form Across Transformer Depth
DBPnet: Damper Characteristics-Based Bayesian Physics-Informed Neural Network for Wheel L…
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning
Adversarial Error Correction for Visual Autoregressive Generation