Halt Fast! Early Stopping for Certified Robustness
Explorar
Noticias de IA
30724 elementos — filtrados, clasificados y sin duplicados
Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes
Agentic Hardware Design as Repository-Level Code Evolution
Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives
Towards Evaluation of Implicit Software World Models in Coding LLMs
Agent-Native Immune System: Architecture, Taxonomy, and Engineering
Tandem Reinforcement Learning with Verifiable Rewards
AI-Driven Synthesis for High-Tech System Design: Automating Innovation
STAG: Spatio-temporal Evolving Structural Representation of Action Units for Micro-expres…
OperatorSHAP: Fast and Accurate Shapley Value Estimation for Neural Operators
Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs
Hippocampus-DETR: An Explicit Memory Object Detection Framework Based on Hippocampus Mode…
Ontology-Guided Evidence Path Inference for Multi-hop Knowledge Graph Question Answering
Verifiable Geometry Problem Solving: Solver-Driven Autoformalization and Theorem Proposing
Towards Reliable and Robust LLM Planning: Symbolic Feedback-Driven Iterative Self-Refinem…
MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy
Mitigating LLM-based p-Hacking by Preregistering for the Next LLM
EAGT: Echocardiography Augmentation for Generalisability and Transferability
Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics …
The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their …
GenMatter: Perceiving Physical Objects with Generative Matter Models
LiveClawBench: Benchmarking LLM Agents on Complex, Real-World Assistant Tasks
From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Sc…
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
Measuring the Redundancy of Decoder Layers in SpeechLLMs
MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
HiMu: Hierarchical Multimodal Frame Selection for Long Video Question Answering
Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees…
When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing …
Spectral Text Fusion: A Frequency-Aware Approach to Multimodal Time-Series Forecasting