Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
P2MFDS: A Privacy-Preserving Multimodal Fall Detection System for Elderly People in Bathr…
Beyond Memory: A Transactional Continuity Kernel for Long-Lived AI Agents
Motion-as-Prompt: Enhancing Motion Reasoning in Multimodal Large Language Models via Moti…
Rubric Dropout: A Simple Way to Mitigate Reward Hacking in Rubric-as-Reward RL
Commonsense on Demand: Generating and Selectively Integrating Commonsense Knowledge for N…
Empowering Children to Create AI-Enabled Augmented Reality Experiences
CORE-3D: Context-aware Open-vocabulary Retrieval by Embeddings in 3D
GCPO: Diagnosing and Constraining Subspace Geometry in Rollout RL for LLMs
MicroAUNet: Boundary-Enhanced Multi-scale Fusion with Knowledge Distillation for Colonosc…
A-3PO: Accelerating Asynchronous LLM Training with Staleness-aware Proximal Policy Approx…
LLM-Powered Automatic Translation and Urgency in Crisis Scenarios
How effective are VLMs in assisting humans in inferring the quality of mental models from…
A Simple Efficiency Incremental Learning Framework via Vision-Language Model with Nonline…
Large Language Models Reproduce Racial Stereotypes When Used for Text Annotation
Forecasting Side Effects of Activation Steering
Cutting AI Datacenter Energy with Reinforcement Learning: Measured Power Control of LLM T…
Designing Agentic AI-Based Screening for Portfolio Investment
Identity from the Outside: A Conceptual Framework and Research Program for AI Personality…
Harnessing agent memory to build lifelong AI partners for materials scientists
AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research
Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Da…
Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonom…
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy F…
Learning from Multimodal Pseudo-Labels for Robust Open-Vocabulary Instance and Panoptic S…
APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference
Locating and Controlling Implicit Personalization in Large Language Models
Pretraining large language models with MXFP4 on Native FP4 Hardware
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
CAR: Query-Guided Confidence-Aware Reranking for Retrieval-Augmented Generation