Browse

AI News

27413 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
Uncertainty-Aware Calibrated Clinical Text Classification with Large Language Models
arXiv cs.AI Research & Papers
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
arXiv cs.AI Research & Papers
TyPatch: Transforming Patches into Typestate Rules for Kernel Bug Detection
arXiv cs.AI Research & Papers
Perceptual Reality Transformer: What Must an Illustration Preserve?
arXiv cs.AI Research & Papers
Efficient On-Device Agents via Adaptive Context Management
arXiv cs.AI Research & Papers
Concertina: Data-Centric Adaptive Pipeline Parallelism for Efficient Heterogeneous Long-C…
arXiv cs.AI Research & Papers
Generalized Correctness Models: Learning Calibrated and Model-Agnostic Correctness Predic…
arXiv cs.AI Research & Papers
CapGeo-Bench: Decoupling Visual Perception from Reasoning and Evaluating Geometric Unders…
arXiv cs.AI Research & Papers
Focus on What Matters: Fisher-Guided Adaptive Multimodal Fusion for Vulnerability Detecti…
arXiv cs.AI Research & Papers
Large Language Models As Shannon Lossy Compressors Not Solomonoff Induction Estimators: T…
arXiv cs.AI Research & Papers
AbFlow : End-to-end Paratope-Centric Antibody Design by Interaction Enhanced Flow Matching
arXiv cs.AI Research & Papers
Learning Human-Like Badminton Skills for Humanoid Robots
arXiv cs.AI Research & Papers
Testing the Limits of Truth Directions in LLMs
arXiv cs.AI Research & Papers
Can We Still Trace L1 Signals? Investigating the Resilience of Native Language Signals in…
arXiv cs.AI Research & Papers
Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Mode…
arXiv cs.AI Research & Papers
Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents
arXiv cs.AI Research & Papers
ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs
arXiv cs.AI Research & Papers
Attention Is All You Need (to Avoid Spurious Oscillations)
arXiv cs.AI Research & Papers
HarnessBandit: Joint Learnability-Transferability Scheduling for Multi-Harness Agentic Re…
arXiv cs.AI Research & Papers
Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Ha…
arXiv cs.AI Research & Papers
Tool Use Reduces Depth-Induced Collapse in OOD Reasoning
arXiv cs.AI Research & Papers
To do($x$) or not to do($x$): Medical Image Counterfactuals for Dataset Augmentation
arXiv cs.AI Research & Papers
Towards Evolving Context Parameterization for Large Language Models
arXiv cs.AI Research & Papers
Same Patient, Different Order: Action-Level Reliability of Clinical LLM Agents Under Repe…
arXiv cs.AI Research & Papers
GEAR: From Dynamic Encoding to Dynamic Activation in Social Trajectory Prediction
arXiv cs.AI Research & Papers
Hindsight Bias in Clinical Temporal Reasoning: How Future Data Exposure Affects Large Lan…
arXiv cs.AI Research & Papers
Hardware-Aware Learned Representation Compression for Distributed In-Sensor Vision
arXiv cs.AI Research & Papers
Thought without systematicity? Evaluating reasoning models on rule induction tasks
arXiv cs.AI Research & Papers
Confuse the Model, Control the Flow: Understanding and Mitigating Privacy Leakage from LL…
arXiv cs.AI Research & Papers
Atria Dawn: The Dawn of Agentic Superintelligence