Uncertainty-Aware Calibrated Clinical Text Classification with Large Language Models
Browse
AI News
27413 items — filtered, classified, deduplicated
LLM Probability Concentration: How Alignment Shrinks the Generative Horizon
TyPatch: Transforming Patches into Typestate Rules for Kernel Bug Detection
Perceptual Reality Transformer: What Must an Illustration Preserve?
Efficient On-Device Agents via Adaptive Context Management
Concertina: Data-Centric Adaptive Pipeline Parallelism for Efficient Heterogeneous Long-C…
Generalized Correctness Models: Learning Calibrated and Model-Agnostic Correctness Predic…
CapGeo-Bench: Decoupling Visual Perception from Reasoning and Evaluating Geometric Unders…
Focus on What Matters: Fisher-Guided Adaptive Multimodal Fusion for Vulnerability Detecti…
Large Language Models As Shannon Lossy Compressors Not Solomonoff Induction Estimators: T…
AbFlow : End-to-end Paratope-Centric Antibody Design by Interaction Enhanced Flow Matching
Learning Human-Like Badminton Skills for Humanoid Robots
Testing the Limits of Truth Directions in LLMs
Can We Still Trace L1 Signals? Investigating the Resilience of Native Language Signals in…
Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Mode…
Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents
ModaLens: Measuring Image Sensitivity in Report-Conditioned Medical VLMs
Attention Is All You Need (to Avoid Spurious Oscillations)
HarnessBandit: Joint Learnability-Transferability Scheduling for Multi-Harness Agentic Re…
Bench2Dex: Benchmarking Visuo-Tactile Bimanual Dexterous Manipulation Across Dexterous Ha…
Tool Use Reduces Depth-Induced Collapse in OOD Reasoning
To do($x$) or not to do($x$): Medical Image Counterfactuals for Dataset Augmentation
Towards Evolving Context Parameterization for Large Language Models
Same Patient, Different Order: Action-Level Reliability of Clinical LLM Agents Under Repe…
GEAR: From Dynamic Encoding to Dynamic Activation in Social Trajectory Prediction
Hindsight Bias in Clinical Temporal Reasoning: How Future Data Exposure Affects Large Lan…
Hardware-Aware Learned Representation Compression for Distributed In-Sensor Vision
Thought without systematicity? Evaluating reasoning models on rule induction tasks
Confuse the Model, Control the Flow: Understanding and Mitigating Privacy Leakage from LL…
Atria Dawn: The Dawn of Agentic Superintelligence