Browse

AI News

26304 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
RL-Lock: Reinforcement Learning for Generating Interlocking Assemblies
arXiv cs.AI Research & Papers
An AI-Based Decision-Support Pipeline for Day-Ahead Photovoltaic Forecasting
arXiv cs.AI Research & Papers
Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning
arXiv cs.AI Research & Papers
Fighting Fire with Fire: On the Feasibility of Protecting Exercises Against AI Cheating
arXiv cs.AI Research & Papers
Bounded Normative Equivalence in Human-AI Cooperation: Group Behaviour, Not Partner Label…
arXiv cs.AI Research & Papers
REFLEX: Rethinking MoE Inference as Refinement-Aware Compute Allocation in Diffusion Lang…
arXiv cs.AI Research & Papers
PB$^2$: Preference Space Exploration via Population-Based Methods in Preference-Based Rei…
arXiv cs.AI Research & Papers
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving V…
arXiv cs.AI Models & Releases
DiffusionGemma Technical Report
arXiv cs.AI Research & Papers
Retrieval Augmented Biomedical Question Answering with Weak Question Recovery and Neural …
arXiv cs.AI Research & Papers
CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning
arXiv cs.AI Research & Papers
Corrigible Assistance in One Round: Pragmatic-Pedagogic Best Response
arXiv cs.AI Research & Papers
Cost-Effective Automated Judging of Natural-Language Mathematical Proofs
arXiv cs.AI Research & Papers
Boosting Generalizable Depth Estimation in Endoscopy by Mixture of Lightweight Experts an…
arXiv cs.AI Research & Papers
MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents
arXiv cs.AI Research & Papers
Cross-Benchmark Generalization in Long-Horizon Agents
arXiv cs.AI Research & Papers
SearchMaster: Grounded and Regulated Self-Play for Search Agents
arXiv cs.AI Research & Papers
Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction
arXiv cs.AI Research & Papers
The Scaling Paradox in Human-AI Collaboration
arXiv cs.AI Research & Papers
It's the Decoding Format, Not the Perturbation: Auditing Consistency-Based Selection for …
arXiv cs.AI Research & Papers
Topology Enhanced MARL for Multi-Agent Cooperative Decision-Making of CAVs
arXiv cs.AI Research & Papers
PATH-Bench: Path-Dependent Evaluation of Lifelong Agents
arXiv cs.AI Research & Papers
Divisive Normalization Shapes Low-Rank Slow Manifolds for Continuous Working Memory
arXiv cs.AI Research & Papers
FAST-GS: Frequency Aware Space-time Gaussian Splatting for Photorealistic Dynamic Novel V…
arXiv cs.AI Research & Papers
Co-evolution of social reward and punishment under institutional interventions
arXiv cs.AI Research & Papers
Hard Constraints, Smooth Gradients: Learning Feasible Inventory Policies via Differentiab…
arXiv cs.AI Research & Papers
Neuro-Symbolic Participation Governance for Verifiable AI Agents in Open Digital Twin Eco…
arXiv cs.AI Research & Papers
Exploring and Bridging Knowledge Holes in Unlearned Multimodal Large Language Models
arXiv cs.AI Research & Papers
Cooperative Coevolution for Resource-Constrained Agentic LLM Post-Training
arXiv cs.AI Research & Papers
MedUPS: Towards Diagnostic Assistance in Uncommon Medical Cases with Large Language Models