RL-Lock: Reinforcement Learning for Generating Interlocking Assemblies
Browse
AI News
26304 items — filtered, classified, deduplicated
An AI-Based Decision-Support Pipeline for Day-Ahead Photovoltaic Forecasting
Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning
Fighting Fire with Fire: On the Feasibility of Protecting Exercises Against AI Cheating
Bounded Normative Equivalence in Human-AI Cooperation: Group Behaviour, Not Partner Label…
REFLEX: Rethinking MoE Inference as Refinement-Aware Compute Allocation in Diffusion Lang…
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Rei…
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving V…
DiffusionGemma Technical Report
Retrieval Augmented Biomedical Question Answering with Weak Question Recovery and Neural …
CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning
Corrigible Assistance in One Round: Pragmatic-Pedagogic Best Response
Cost-Effective Automated Judging of Natural-Language Mathematical Proofs
Boosting Generalizable Depth Estimation in Endoscopy by Mixture of Lightweight Experts an…
MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents
Cross-Benchmark Generalization in Long-Horizon Agents
SearchMaster: Grounded and Regulated Self-Play for Search Agents
Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction
The Scaling Paradox in Human-AI Collaboration
It's the Decoding Format, Not the Perturbation: Auditing Consistency-Based Selection for …
Topology Enhanced MARL for Multi-Agent Cooperative Decision-Making of CAVs
PATH-Bench: Path-Dependent Evaluation of Lifelong Agents
Divisive Normalization Shapes Low-Rank Slow Manifolds for Continuous Working Memory
FAST-GS: Frequency Aware Space-time Gaussian Splatting for Photorealistic Dynamic Novel V…
Co-evolution of social reward and punishment under institutional interventions
Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiab…
Neuro-Symbolic Participation Governance for Verifiable AI Agents in Open Digital Twin Eco…
Exploring and Bridging Knowledge Holes in Unlearned Multimodal Large Language Models
Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training
MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models