FIGMA: Towards FIne-Grained Music retrievAl
Explorar
Noticias de IA
22065 elementos — filtrados, clasificados y sin duplicados
ChronoForest: Closed-Loop Multi-Tree Diffusion Planning for Efficient Bridge Search and R…
What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos?
Inside the Visual Mind: Neuroscience-Motivated Concept Circuits for Interpreting and Stee…
HKJudge: A Legal Discourse-Annotated Corpus for Interpreting What Courts Find, How They R…
The Geography of Algorithmic Judgment: LLM Intermediaries, Place Identity, and Racial Ste…
Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation
ShallowBench: Benchmarking Generative Drug Design Models on Shallow-Pocket Targets
MSAIC-Net: A Multi-Scale Attention and Imbalance-Aware Contrastive Network for ECG-Based …
SCOUT: Semantic scene COverage via Uncertainty-guided Traversal
HybridCodec: Fast Dual-Stream, Semantically Enhanced Neural Audio Codec
Evidence Graph Consistency in Retrieval-Augmented Generation: A Model-Dependent Analysis …
AxisGuide: Grounding Robot Action Coordinate System in RGB Observations for Robust Visuom…
Optimal Rates for Generalization of Gradient Descent Methods with Deep Neural Networks
Mind the Gap: Bridging Behavioral Silos with LLMs in Multi-Vertical Recommendations
Exploring Reinforcement Learning for Fluid Transitions Between Clinical Mental Healthcare…
Breaking the Lock-in: Diversifying Text-to-Image Generation via Representation Modulation
PandaAI: A Practical Agent CQ2 for Neuro-symbolic Data Analysis And Integrated Decision-M…
Hearing the Unspoken: Language Model Priors for Acoustic Adversarial Attacks
LLM Agent-Assisted Reverse Engineering with Quantitative Readability Metrics
Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces
MotionEnhancer: Leveraging Video Diffusion for Motion-Enhanced Vision-Language Models
Modeling Nonlinear Feature Interactions with Product-Unit Residual Networks
EgoPressDiff: Multimodal Video Diffusion for Egocentric UV-Domain Hand-Pressure Estimation
Neuro-Symbolic Learning for Long-Horizon Task Planning Under Complex Logical Constraints
EASE-TTT: Evidence-Aligned Selective Test-Time Training for Long-Context Question Answeri…
ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning
Auditing Training Data in Domain-adapted LLMs: LoRA-MINT
OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Sce…
DaX: Learning General Pathology Representations Across Scales