Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preferen…
Explorar
Noticias de IA
29670 elementos — filtrados, clasificados y sin duplicados
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
Anatomy-Guided Vision-Language Learning with Angular Prototype Separation for Multi-Label…
Safe Reinforcement Learning with Preference-based Constraint Inference
Design and Report Benchmarks for Knowledge Work
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide …
HTMuon: Improving Muon via Heavy-Tailed Spectral Correction
Understanding Goal Generalisation in Sequential Reinforcement Learning
ETCHR: Editing To Clarify and Harness Reasoning
Visually-Guided Policy Optimization for Multimodal Reasoning
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transform…
CHRONOS: Temporally-Aware Multi-Agent Coordination for Evolving Data Marketplaces
Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPO
Any2Any: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking
GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models
PhotoFlow: Agentic 3D Virtual Photography Missions
Worse than Random: The Importance of a Baseline for Unsupervised Feature Selection
Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good
SUDP: Secret-Use Delegation Protocol for Agentic Systems
SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety
GenAI-Driven Threat Detection with Microsoft Security Copilot
OnePred: Next-Query Prediction via Recursive Intent Memory in Multi-Turn Conversations
It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amp…
Redrawing the AI Map: A Theory of Accountability Boundaries in Agentic Ecosystems
The AI-Native Large-Scale Agile Software Development Manifesto
HARNESS-LM: A Three-Phase Training Recipe for Harnessing SLMs in Sponsored Search Retriev…
EMA-Nesterov: Stabilizing Nesterov's Lookahead for Accelerated Deep Learning Optimization
Apple Watchに変革を、Whoopやオーラ台頭でヘルスケアアプリに課題-Power On
Sakura Internet Eyes More Spending to Meet Japan’s AI Demand
Learning to Route Languages for Multilingual Policy Optimization