RRPO: Reference-Relative Policy Optimization with Stratified Conditional Rollouts
Explorar
Noticias de IA
22116 elementos — filtrados, clasificados y sin duplicados
LatentMT: Machine Translation with Latent Reasoning
Mitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learni…
Calibrated Selective Fact-Checking via Evidence Chain Evaluation
Competitive and Complementary Tools
AutoIndex: Learning Representation Programs for Retrieval
When JSON Is Not Enough: Semantic Reliability of Schema-Constrained LLM Ordering Agents
Athena-Brain Technical Report: An Efficient Robot Brain for General Intelligence and Embo…
On the Effectiveness of Pretraining for Graph Combinatorial Optimization
AI Tool Discovery at Scale: All You Need is DNS
Memo2496: Expert-Annotated Dataset and Dual-view Adaptive Framework for Music Emotion Rec…
Position: AI/ML Deepfake Research is Misaligned with AI-Generated Non-Consensual Intimate…
Structured Synthetic Reasoning Data for Arithmetic Fine-Tuning of Small Language Models
ChainMark: Model-Free LLM Watermarking with Closed-Form Calibration
Comparative Study of Multi-Agent Actor-Critic Algorithms in Parameterized Action Reinforc…
Sequential Learner Modeling Using Multi-Relational Graph Convolutional Networks
Federated Lightweight Fine-Tuning
Relay-Bench: Evaluating LLMs on Multi-Domain Reasoning Chains
ISO: An RLVR-Native Optimization Stack
ChemHyperMag: Physics-informed magnetic hypergraph learning improves molecular ADMET pred…
EvalSafetyGap: A Hybrid Survey and Conceptual Framework for LLM Evaluation-Safety Failures
Associative Emotional Learning in Convolutional Neural Networks
Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs
Agentic AI-assisted coding offers a unique opportunity to instill epistemic grounding dur…
Analytic Distribution of Classifier-Free Guidance for Schedule Design
An Exploratory Analysis of Pain Localization via Explainable Computational Modeling
ReFace: Reorganizing Facial Spatiotemporal Representations for Improved Pain Assessment
How Fast Can Reward Models Score? A Systems Study of C++ and PyTorch Inference Runtimes f…
PhenSPINE: A Standardized Benchmark for Spine Pathology Diagnosis
NavVerse: Benchmarking Indoor-to-Outdoor Embodied Navigation in Continuous Robot Simulati…