Browse

AI News

30934 items — filtered, classified, deduplicated

arXiv cs.AI Developer Tooling
MinT: Managed Infrastructure for Training and Serving Millions of LLMs
arXiv cs.AI Research & Papers
AgentSociety: Incentivizing Agentic Social Intelligence
arXiv cs.AI Security & Safety
Black-box Membership Inference Attacks on the Pre-training Data of Image-generation Models
arXiv cs.AI Research & Papers
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
arXiv cs.AI Research & Papers
LR-SGS: Robust LiDAR-Reflectance-Guided Salient Gaussian Splatting for Self-Driving Scene…
arXiv cs.AI Research & Papers
Unified Neural Scaling Laws
arXiv cs.AI Research & Papers
Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning
arXiv cs.AI Research & Papers
ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Fe…
arXiv cs.AI Research & Papers
Identifiable Token Correspondence for World Models
arXiv cs.AI Research & Papers
VesselSim: learning 3D blood vessel segmentation without expert annotations
arXiv cs.AI Research & Papers
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
arXiv cs.AI Research & Papers
Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation ac…
arXiv cs.AI Research & Papers
ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis
arXiv cs.AI Research & Papers
Practical Anonymous Two-Party Gradient Boosting Decision Tree
arXiv cs.AI Research & Papers
ICICLE: Expanding Retrieval with In-Context Documents
arXiv cs.AI Research & Papers
EHRSummarizer: A Privacy-Aware, FHIR-Native Reference Architecture for Source-Grounded EH…
arXiv cs.AI Research & Papers
Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents
arXiv cs.AI Research & Papers
GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations
arXiv cs.AI Research & Papers
ScenePilot: Controllable Boundary-Driven Critical Scenario Generation for Autonomous Driv…
arXiv cs.AI Research & Papers
AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito
arXiv cs.AI Research & Papers
Neuro-Symbolic Verification of LLM Outputs for Data-Sensitive Domains (extended preprint)
arXiv cs.AI Research & Papers
CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generations
arXiv cs.AI Research & Papers
Beyond a Single Direction: Chain-of-Thought Disrupts Simple Steering of Refusal
arXiv cs.AI Research & Papers
ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis
arXiv cs.AI Research & Papers
Auditing and Fixing Economic Validity in Tabular Foundation Models for Discrete Choice
arXiv cs.AI Research & Papers
Decoupled Delay Compensation: Enhancing Pre-trained MARL Policies via Learned Dynamics Fi…
arXiv cs.AI Research & Papers
GeoFaith: A Spatio-Temporal Dual View of Faithful Chain-of-Thought
arXiv cs.AI Research & Papers
Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations
arXiv cs.AI Research & Papers
Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language …
arXiv cs.AI Research & Papers
The Sensation Modulating Network:Haltability as the architectural ground for object-direc…