Browse

AI News

25879 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
Leveraging AI for fine-grained food safety risk forecasting in sparse data conditions
arXiv cs.AI Research & Papers
Diagnosing Search Behavior and Failure Modes in Long-Horizon Search Agents
arXiv cs.AI Research & Papers
Neuro-Evolved Heuristics for Variable Gapped Common Subsequence Identification
arXiv cs.AI Research & Papers
RL-Lock: Reinforcement Learning for Generating Interlocking Assemblies
arXiv cs.AI Research & Papers
Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation
arXiv cs.AI Research & Papers
When Memory Updates but Behavior Does Not: Repairing Implicit Stale Dependencies in Perso…
arXiv cs.AI Research & Papers
EduZone: A Framework for Evaluating LLM Safety for K-12 Students and Teachers
arXiv cs.AI Research & Papers
Tracing the Cascade: A Topology-Aware Evaluation Framework for Scientific Agent Hallucina…
arXiv cs.AI Research & Papers
AI-Based Thesis Assessment: An Empirical Study of Human Evaluation Priorities and Their I…
arXiv cs.AI Research & Papers
AgentSLABench: Evaluating and Benchmarking Agentic Systems Under Resource Constraints
arXiv cs.AI Research & Papers
CADIR: A Cross-Backend Editable Intermediate Representation for Agentic CAD Generation
arXiv cs.AI Research & Papers
GISAgentBench: A Practitioner-Sourced Benchmark for Evaluating LLM Agents on GIS Tasks
arXiv cs.AI Research & Papers
Fetch-then-Explore: Decoupling Selection from Extraction over a Persistent Workspace for …
arXiv cs.AI Research & Papers
Beyond the Mean: Multi-Moment Policy Optimization for LLM Reasoning
arXiv cs.AI Research & Papers
DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimizat…
arXiv cs.AI Research & Papers
Can You Trust the Confidence? ConfBench for Vision-Language Models on Document Extraction
arXiv cs.AI Research & Papers
Slides2MindMap: Reconstructing Cognitively Efficient Knowledge Hierarchies from Lecture S…
arXiv cs.AI Research & Papers
TrajWiki: Source-Grounded Memory Trajectories for Long-Horizon Dialogue Agents
arXiv cs.AI Research & Papers
CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning
arXiv cs.AI Research & Papers
Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Mode…
arXiv cs.AI Research & Papers
HetGPS: Scalable Graph Multi-Agent Reinforcement Learning with Physics-Anchored Adaptive …
arXiv cs.AI Research & Papers
EchoChange: A Diffusion Language Model with Dual Pass Remasking for Factual Remote Sensin…
arXiv cs.AI Research & Papers
CURE: Local Uncertainty Repair for Block-Parallel Speculative Decoding
arXiv cs.AI Research & Papers
Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations
arXiv cs.AI Research & Papers
Multi-Dimensional Assessment for AI Cognition (MAAC): A Theoretical Framework for Process…
arXiv cs.AI Research & Papers
BayesSeg: A Bayesian Optimization Framework for State Segmentation of Electricity Consump…
arXiv cs.AI Research & Papers
The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence
arXiv cs.AI Research & Papers
Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models
arXiv cs.AI Research & Papers
CockpitHAT: Dependency-Graph-Driven Hierarchical Attribution for Embodied Multi-Agent Coc…
arXiv cs.AI Research & Papers
Evolving in the Agent Jungle via History-Informed Opponent Awareness