A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integ…
Explorar
Noticias de IA
29349 elementos — filtrados, clasificados y sin duplicados
Bi-Level Reinforcement Learning Pathway for Sim-to-Real Optimality
When Large Language Models Know the Table: A Framework for Assessing Data Contamination i…
Neural Diversity Regularizes Hallucinations in Language Models
Interpreting GFlowNets for Drug Discovery: What probes can and cannot show
Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tok…
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
Zero-shot reasoning for simulating scholarly peer-review
Chained Recursive Language Models for Multi-Iteration Reasoning
OPD-V: Visual On-Policy Self-Distillation with Modality Balance
GRALS: GCN-Guided Redundancy-Aware Local Search for Minimum Vertex Cover
Large-Small Model Collaboration for Enhancing Edge-Deployed Small Models
Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-…
Curiosity-Diffuser: Curiosity Guide Diffusion Models for Reliability
Capability-Gated Planning: Cost-to-Goal Discovery and the Limits of Myopic Experiment Sel…
Representational separation between unitary and channel quantum generative models via sha…
Arnold: A multi-task, multi-embodiment muscle transformer policy
ArtAnno: Annotating Implicit Semantics in Artworks through LLM Agent-Driven Bidirectional…
OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents
Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning
SVI-DAG: A Structured Variational Inference Approach to Bayesian Causal Discovery
The Effect of Perceived Race and Gender on Police Language Use: Experimental Evidence fro…
Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation
A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination
RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists
A Chain Is Only as Strong as Its Weakest Link: A Scoping Review of System Integration Aud…
Towards a satellite image manipulation and deepfake localization benchmark dataset
Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning
RepairFormer: Automated Repair of Structured Inputs Using Transformers