Browse

AI News

30934 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations
arXiv cs.AI Research & Papers
Recursive Flow Matching
arXiv cs.AI Research & Papers
Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning
arXiv cs.AI Research & Papers
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments
arXiv cs.AI Research & Papers
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentat…
arXiv cs.AI Research & Papers
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
arXiv cs.AI Research & Papers
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
arXiv cs.AI Research & Papers
The ATOM Report: Measuring the Open Language Model Ecosystem
arXiv cs.AI Research & Papers
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-…
arXiv cs.AI Research & Papers
SenBen: Sensitive Scene Graphs for Explainable Content Moderation
arXiv cs.AI Security & Safety
Cryptographic Registry Provenance: Structural Defense Against Dependency Confusion in AI …
arXiv cs.AI Research & Papers
Understanding the Challenges in Iterative Generative Optimization with LLMs
arXiv cs.AI Research & Papers
From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Ques…
arXiv cs.AI Research & Papers
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argum…
arXiv cs.AI Research & Papers
StreamSplit: Continuous Audio Representation Learning via Uncertainty-Guided Adaptive Spl…
arXiv cs.AI Policy & Regulation
Algorithmic Monocultures in Hiring
arXiv cs.AI Research & Papers
Detached Skip-Links and $R$-Probe: Decoupling Feature Aggregation from Gradient Propagati…
arXiv cs.AI Research & Papers
Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinfo…
arXiv cs.AI Research & Papers
APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented…
arXiv cs.AI Research & Papers
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
arXiv cs.AI Research & Papers
FedTreeLoRA: Reconciling Statistical and Functional Heterogeneity in Federated LoRA Fine-…
arXiv cs.AI Research & Papers
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
arXiv cs.AI Research & Papers
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
arXiv cs.AI Research & Papers
Geometrically Constrained Outlier Synthesis
arXiv cs.AI Research & Papers
GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation
arXiv cs.AI Research & Papers
Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User Hist…
arXiv cs.AI Research & Papers
Adapting Actively on the Fly: Relevance-Guided Online Meta-Learning with Latent Concepts …
arXiv cs.AI Research & Papers
Constructing Industrial-Scale Optimization Modeling Benchmark
arXiv cs.AI Research & Papers
GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Pos…
arXiv cs.AI Research & Papers
Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Sp…