Upper Bounds on the Generalization Error of Deep Learning Models via Local Robustness and…
Explorar
Noticias de IA
30288 elementos — filtrados, clasificados y sin duplicados
Compositional Reasoning Depth Predicts Clinical AI Failure: Empirical Evidence Consistent…
Semantic Flip: Synthetic OOD Generation for Robust Refusal in Embodied Question Answering…
Binary Tracking for Spatial QA and Navigation with Open Vision-Language Models
Scalable Circuit Learning for Interpreting Large Language Models
Phantoms and Disclosures: a Causal Framework for Auditing Synthetic Data
ActiveSAM: Image-Conditional Class Pruning for Fast and Accurate Open-Vocabulary Segmenta…
HAMON: Passive Optical Sequence Mixing for Long-Horizon Forecasting
Attention, not scale, drives human-AI alignment in multimodal language prediction
Multi-Sensor Fusion for UAV Classification Based on Feature Maps of Image and Radar Data
Computational Safety for Generative AI: A Hypothesis Testing Perspective
Fine-Tuning a 7B Advisor on Free-Tier GPUs: An Adapter-Handoff Recipe and a Synthetic-Dat…
Unifying Post-hoc Explanations of Knowledge Graph Completions
Optimizing Health Coverage in Ethiopia: A Learning-augmented Approach and Persistent Prop…
Shachi: A Modular, Controllable Framework for LLM-Based Agent-Based Modeling of Emergent …
NeuronFabric: A Software Reference Architecture for On-Chip Transformer Training with Loc…
Sample from What You See: Visuomotor Policy Learning via Diffusion Bridge with Observatio…
Multi-Granular Node Pruning for Causal Circuit Discovery
MedAI: Evaluating TxAgent's Therapeutic Agentic Reasoning in the NeurIPS CURE-Bench Compe…
E-mem: Multi-agent based Episodic Context Reconstruction for LLM Agent Memory
AgentLeak: A Benchmark for Internal-Channel Privacy Leakage in Multi-Agent LLM Systems
SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
A Model-Free Universal AI
SorryDB: Can AI Provers Complete Real-World Lean Theorems?
EMS: Multi-Agent Voting via Efficient Majority-then-Stopping
Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalize…
The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling
FORTIS: Benchmarking Over-Privilege in Agent Skills
A Survey on 3D Skeleton Based Person Re-Identification: Taxonomy, Advances, Challenges, a…
Canonical Variates in Wasserstein Metric Space