TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Vid…
Explorar
Noticias de IA
30239 elementos — filtrados, clasificados y sin duplicados
Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length …
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
Separating Secrets from Placeholders: A Hybrid CNN-CodeBERT Framework for Three-Class Cre…
On Efficient Scaling of GNNs via IO-Aware Layers Implementations
Fine-grained Verification via Diagnostic Reasoning Supervision for Aspect Sentiment Tripl…
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Info…
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation w…
The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of La…
Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Mod…
Scaling Higher-Order Graph Learning with Maximal Clique Complexes
dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Deve…
Appropriateness of Empathy in AI: A Signal-Cost Perspective
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decisio…
SAM for Robust Mitochondria Instance Segmentation in Fluorescence Microscopy
Practical Cross-Band Channel Prediction for AI-RAN via Physics-Guided Deep Unfolding
Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Un…
Correcting Split Selection in Online Decision Trees via Anytime-Valid Inference
What changes after deployment? A survey on On-device Learning in TinyML
Benchmarking and Enhancing Text-to-Image Models for Generating Visual Representations in …
Towards Atoms of Large Language Models
A Kinetic Energy Perspective of Flow Matching
Probing Collision Grounding in Vision-Language Models for Safe Human-Robot Collaboration
MindVoice: Reconstructing Intelligible Speech from Non-invasive Neural Signals with Pretr…
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Over…
Trust-Region Behavior Blending for On-Policy Distillation
Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phr…
Developing a Culturally Grounded, AI-Augmented UX Research Point of View (POV): An Exempl…
SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes