Exploring and Bridging Knowledge Holes in Unlearned Multimodal Large Language Models
Browse
AI News
26304 items — filtered, classified, deduplicated
When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling
From fragmented data to actionable design: Physics-calibrated learning for plastic upcycl…
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models
Latent Thought Credit: Multi-Answer Credit Assignment for Latent Reasoning
Human-Centered Reflections on Care Robots: A Comparative Study of Caregiver Perspectives
SIEVE: Selective Integrity Verification and Escalation for Defending LLM Agents against I…
InfoOps Bench: A live information operations safety benchmark
Meritocratic Fairness via $K$-Shapley Values in Budgeted Combinatorial Bandits with Full-…
Counterfactual Reasoning for Causal Responsibility Attribution in Probabilistic Multi-Age…
Context-Aware Mixture of Domain Experts for Bodily Expression of Emotion in the Wild
BRiG-AFA: Bellman Risk-to-Go Learning for Non-Myopic Active Feature Acquisition
GROVE: Growing and Reasoning over Temporally Stratified Memory from Streaming Video Exper…
Diffusion Policy with Behavioral Advantage Correction for Offline Reinforcement Learning
Can AI Agents Simulate A/B Test Outcomes? A Validation Framework for Agentic Experimentat…
The Elements of Differentiable Programming
HarMoE: Multi-Source Chest Radiograph Pretraining with Dataset-Disentangled Experts
Can Foundation Models Hear What Made That Sound? A Tiered Benchmark of Audio-Language Mod…
PolymerGPT: Multi-property Optimization with a Decoder-Based GPT Model for Generative Pol…
TS-MAMP: A Remanufactured Agricultural Robot Powered by Second-Life EV Components and NMS…
AI and Its Impact on Creativity and Diversity: An Empirical Study of LLM-Generated Produc…
Understanding Machine Unlearning Through the Lens of Mode Connectivity
Where did the ambiguity go? Examining how multimodal models interpret polysemous words
UniqueSplat: View-conditioned 3D Gaussian Splatting for Generalizable 3D Reconstruction
PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-…
IACM-RL: Intent-Aware Context Management and Reinforcement Learning for Complex Tool Invo…
Securing Agentic AI: From Per-Action Checks to Trajectory Assurance
Diagnose Before You Compress: Prediction-Independent Bottleneck Witness Refinement for LL…
Self-Improving Large Language Models via Progressive Experience Evolution
Assessing the Impacts of Imperfect Datasets on Client Selections in Federated Learning