Risk-Aware Selective Multimodal Driver Monitoring with Driver-State World Modeling
Explorar
Noticias de IA
30711 elementos — filtrados, clasificados y sin duplicados
Bridging Vision and Language Concepts through Optimal Transport Semantic Flow
NaviCache: Test-Time Self-Calibration Caching for Video Generation
ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP
MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG
Anatomy-Guided Residual Motion Diffusion for Controllable 4D Cardiac MRI Synthesis
MLFFM-SegDiff: A Multi-Level Feature Fusion Diffusion Model for Skin Lesion Segmentation
Algorithmic Foundations of Deep Learning: Complexity-Theoretic Rates and a Characterizati…
Beyond Logical Forms: LLM-Extracted Patterns for Fallacy Classification
Disco-LoRA: Disentangled Composition of Content, Style, and Motion for Multi-concept Vide…
TGHE: Template-based Graph Homomorphic Encryption for Privacy-Preserving GNN Inference in…
CAT-Q: Cost-efficient and Accurate Ternary Quantization for LLMs
Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents
HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visual…
scBench-Long: Verifiable Benchmarking of Long-Horizon Single-Cell Biology
Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI…
CascadeFormer: Depth-Tapered Transformers Motivated by Gradient Fan-in Asymmetry
From Hallucination to Grounding: Diagnosing Visual Spatial Intelligence via CRISP
VoiceTTA: Enhancing Zero-Shot Text-to-Speech via Reinforcement Learning-Based Test-Time A…
An Empirical Study of LLM-Generated Specifications for VeriFast
Speaking Numbers to LLMs: Multi-Wavelet Number Embeddings for Time Series Forecasting
Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents
Retrieval-Warmed Energy-Based Reasoning: A Five-Arm Ablation Methodology for Diffusion-as…
Active Adversarial Perturbation-driven Associative Memory Retrieval for RGB-Event Visual …
ProvenAI: Provenance-Native Traces of Evidence in Generated Answers
Assert, don't describe: Linguistic features that shift LLM reasoning about animal welfare
WatchAct: A Benchmark for Behavior-Grounded Robot Manipulation
Investigating LLM's Problem Solving Capability -- a Study on Statics Questions
Play2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly?
Dream machine -- the next creative economy