Browse

AI News

27413 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
Evaluation of MLLM-Agnostic Plug-and-Play Keyframe Selection Methods for Long Video Under…
arXiv cs.AI Research & Papers
(How) Do MLLMs Report Bistable Images Like Humans?
arXiv cs.AI Research & Papers
From Process Loss to Assembly Bonus: Human-Grounded Diagnosis of Multi-Agent LLM Collabor…
arXiv cs.AI Research & Papers
A Conservative OCR-Enabled Workflow for R214 Sodium Screening of South African Packaged F…
arXiv cs.AI Research & Papers
Multimodal-Multiresolution Foundation Model for Lunar Remote Sensing
arXiv cs.AI Research & Papers
LLaDA-UI: Bringing Block-wise Diffusion to Vision-Language GUI Agents
arXiv cs.AI Research & Papers
Confuse the Model, Control the Flow: Understanding and Mitigating Privacy Leakage from LL…
arXiv cs.AI Research & Papers
From Voice to Value: Leveraging AI to Enhance Spoken Online Reviews on the Go
arXiv cs.AI Research & Papers
SkillAtlas: An Attack Trace Library for Agent Skills
arXiv cs.AI Research & Papers
Multi-Agent Empowerment and Emergence of Complex Behavior in Groups
arXiv cs.AI Research & Papers
Self-Evolving Memory for Generative Recommendation
arXiv cs.AI Research & Papers
Predictive audio representations for early detection and tracking of hidden dynamic objec…
arXiv cs.AI Research & Papers
Through the Eyes of the Beholder: Biometric and Demographic Conditioning for Multimodal S…
arXiv cs.AI Research & Papers
Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering
arXiv cs.AI Research & Papers
Rethinking the Implications of Human Feedback for Preference Learning in Human-Robot Coll…
arXiv cs.AI Research & Papers
Zonal RL-RRT: Integrated RL-RRT Path Planning with Collision Probability and Zone Connect…
arXiv cs.AI Research & Papers
Anatomical Grounding and Leakage-Aware Multimodal Contrastive Learning for Alzheimer's Di…
arXiv cs.AI Research & Papers
SlipSense: Multimodal Tactile Learning for Low-Latency and Generalized Slip Detection
arXiv cs.AI Research & Papers
Mizan: A National Benchmark for Evaluating Large Language Models on Iraqi Arabic and the …
arXiv cs.AI Research & Papers
Hidden APIs in Language Models: Discovering Reusable Causal Interfaces from Forked Futures
arXiv cs.AI Research & Papers
Hardware-Aware Learned Representation Compression for Distributed In-Sensor Vision
arXiv cs.AI Research & Papers
Thought without systematicity? Evaluating reasoning models on rule induction tasks
arXiv cs.AI Research & Papers
From Visual Feedback to Textual Reviews: A Multi-Agent Vision-Language Framework for Imag…
arXiv cs.AI Research & Papers
BGM2Pose: Active 3D Human Pose Estimation with Non-Stationary Sounds
arXiv cs.AI Research & Papers
Adaptive Phase-Switching for Communication-Efficient Federated LoRA Fine-Tuning
arXiv cs.AI Research & Papers
LIMBO: Lifelong Inference-Time Memory and Budget Optimization for LLM Agents
arXiv cs.AI Research & Papers
Talking to Me or Someone Else? Rethinking Talk-to-Me Detection in Egocentric Videos
arXiv cs.AI Research & Papers
Utility-Guided Agent Orchestration for Efficient LLM Tool Use
arXiv cs.AI Research & Papers
RA-CoA: Training-free Fashion Image Captioning via Retrieval-Augmented Chain-of-Attributes
arXiv cs.AI Research & Papers
Safety as a Constraint: Fine-Tuning a LLM Recommender to Explain Itself