Reinforcement Learning for Flow-Matching Policies with Density Transport
Explorar
Noticias de IA
29636 elementos — filtrados, clasificados y sin duplicados
Auditable Graph-Guided Root Cause Analysis for Kubernetes Incidents
Frequency-Domain Latent Attention Gating for Cross-Domain Token Aggregation
Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing
Calibration of Structured Ignorance Certificates for Diagnosing Unknown Unknowns in Reaso…
The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective …
From Statute to Control Flow: Span-Grounded Deontic Trees for Defeasible Scope Parsing
EinSort: Sorting is All We Need for Tensorizing LLM
Closing the Sim-to-Real Gap: An Evaluation Framework for Autonomous Cyber Defense Configu…
When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manip…
ActProbe: Action-Space Probe for Early Failure Detection of Generative Robot Policies
Unambiguous Representations in Neural Networks: An Information-Theoretic Approach to Inte…
Intelligent Character Recognition of Handwritten Forms with Deep Neural Networks
Explaining Data Mixing Scaling Laws
sGPO: Trading Inference FLOPs for Training Efficiency in RLVR
Hybrid Robustness Verification for Spatio-Temporal Neural Networks
Stage-1 Controls the Entropy Regime, Not the Outcome
STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for …
PIPE-Cypher: Automatic Enterprise Benchmark Generation for Text-to-Cypher Systems
FADTI: Fourier and Attention Driven Diffusion for Multivariate Time Series Imputation
Exploring the Effect of Basis Rotation on NQS Performance
Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading…
Evaluating AI Investment Strategies
Supracompetitive Pricing Under AI Monoculture
Revisiting the shutdown problem
PolyBuild: An End-to-End Method for Polygonal Building Contour Extraction from High-Resol…
Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops
RadOT-Eval: Auditable Structured-Evidence Transport for Radiology Report Evaluation
LogNEO: A GPT-Neo Reinforcement Learning Framework for Accurate Real-Time Log Anomaly Det…
Adaptive Loss Balancing for Noise-Robust GRPO in Generative Recommendation