Browse

AI News

18308 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
Where did the ambiguity go? Examining how multimodal models interpret polysemous words
arXiv cs.AI Research & Papers
Bayesian and Motivated Reasoning in AI Agents
arXiv cs.AI Research & Papers
AI-Based Thesis Assessment: An Empirical Study of Human Evaluation Priorities and Their I…
arXiv cs.AI Research & Papers
Neuro-Evolved Heuristics for Variable Gapped Common Subsequence Identification
arXiv cs.AI Research & Papers
SIRIN: A Unified Toolkit for Detecting Contextual Hallucinations in Retrieval-Augmented a…
arXiv cs.AI Research & Papers
Diagnose Before You Compress: Prediction-Independent Bottleneck Witness Refinement for LL…
arXiv cs.AI Research & Papers
TrAC: Trace-Conditioned Answer Consistency for Efficient Uncertainty Quantification in LL…
arXiv cs.AI Research & Papers
F-WANDA: Fisher-Reweighted Post-Training Pruning for Sustainable Deployment of Large Lang…
arXiv cs.AI Research & Papers
Tracing the Cascade: A Topology-Aware Evaluation Framework for Scientific Agent Hallucina…
arXiv cs.AI Research & Papers
Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation
arXiv cs.AI Research & Papers
Modeling Social Dynamics with an LLM-Enabled Agent Based Network-Dynamic (LAND) Model
arXiv cs.AI Research & Papers
More Debate, Same Evidence: Structural Limits of Homogeneous Multi-Agent Groundedness
arXiv cs.AI Research & Papers
Memory Reward Inflation in Self-Improving LLM Agents
arXiv cs.AI Research & Papers
RMSWeb: Reflection, Failure-Mode Mining, and Salvage-DS for Web Agent Reinforcement Learn…
arXiv cs.AI Research & Papers
Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale
arXiv cs.AI Research & Papers
RF-HOI: Recognize Human-Object Interaction with Radio Frequency Signals
arXiv cs.AI Research & Papers
Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations
arXiv cs.AI Research & Papers
AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?
arXiv cs.AI Research & Papers
H+ Embedding: Harmonizing Global and Token-Level Retrieval with Context-Dependent Phrases
arXiv cs.AI Research & Papers
Geometric Self-Supervised Pre-training for Neural Combinatorial Optimization
arXiv cs.AI Research & Papers
Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates
arXiv cs.AI Research & Papers
Motif-Mamba: network motif improved mamba for long-range sequence modeling
arXiv cs.AI Research & Papers
TRACE-TS: Attribution-Grounded and Traceable Sensor-Language Reasoning for Human Activity…
arXiv cs.AI Research & Papers
Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Mode…
arXiv cs.AI Research & Papers
HetGPS: Scalable Graph Multi-Agent Reinforcement Learning with Physics-Anchored Adaptive …
arXiv cs.AI Research & Papers
The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence
arXiv cs.AI Research & Papers
DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimizat…
arXiv cs.AI Research & Papers
Energy Efficiency of Locally Deployed LLMs: A Preliminary Quantitative GPU Power Benchmar…
arXiv cs.AI Research & Papers
AgentSLABench: Evaluating and Benchmarking Agentic Systems Under Resource Constraints
arXiv cs.AI Research & Papers
Assuming You Knew: Fixing an Epistemic Semantics for Flow Policies Using Agentic AI