DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection
Browse
AI News
30934 items — filtered, classified, deduplicated
How Many Tools Should an LLM Agent See? A Chance-Corrected Answer
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorit…
Catching MRI outliers: unsupervised detection and localization of MRI artefacts and clini…
AI-Driven Adaptive Adversaries and the Erosion of Cryptographic Trust in Public Key Syste…
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajecto…
TRAFA: Anticipating User Actions to Reduce Errors in Procedural Tasks with Predictive Fee…
Polymorphism Is Rotation: Operational Mechanistic Interpretability from a Two-Layer Trans…
Batch Normalization Amplifies Memorization and Privacy Risks
Momentum Streams for Optimizer-Inspired Transformers
MX-SAFE: Versatile Inference- and Training-Proof Microscaling Format with On-the-Fly Expo…
ChaosBench-Logic v2: Evaluating LLM Logical Reasoning over Dynamical Systems at Scale
Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malw…
Improving Labeling Consistency with Detailed Constitutional Definitions and AI-Driven Eva…
Action with Visual Primitives
Ordering Matters: Rank-Aware Selective Fusion for Blended Emotion Recognition
FineBench: Benchmarking and Enhancing Vision-Language Models for Fine-grained Human Activ…
ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop
How Few-Shot Examples Add Up: A Causal Decomposition of Function Vectors in In-Context Le…
Continual Speaker Identity Unlearning with Minimal Interference
EchoPilot: Training-Free Ultrasound Video Segmentation via Scale-Space Semantic Prompting…
Small Models, Strong Priors: Architectural Inductive Bias for Parameter-Efficient Neural …
Causal Tongue-Tie: LLMs Can Encode Causal Direction, But Their Yes/No Outputs Fail to Exp…
Context-Instrumental Data Distillation for Kubernetes Manifest Generation: Method and Exp…
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Man…
vAttention: Verified Sparse Attention
Local MAP Sampling for Diffusion Models
Beyond Predefined Learning Objects: A Thinking-Learning Interaction Model for Up-to-Date …
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model