Ouvia: A User-centered Framework for Measuring Usability of Speech Translation in Real-Wo…
Explorar
Noticias de IA
21863 elementos — filtrados, clasificados y sin duplicados
On the training of physics-informed neural operators for solving parametric partial diffe…
$p$-adic Bi-Filtrations for Topological Machine Learning on Genomic Sequences
Where, What, Why, and Importance: Structured Defect Grounding for Text-to-Image Feedback
MS-DKC: A Dataset Knowledge Card Framework for Designing and Adapting Medical Image Segme…
CogManip: Benchmarking Manipulative Behavior in Multi-Turn Interactions with Large Langua…
Learning Visual Spatial Planning from Symbolic State via Modality-Gap-Aware Self-Distilla…
Diffusion Models for Adaptive Sequential Data Generation
ATT-CR: Adaptive Triangular Transformer for Cloud Removal
Fast and Robust Convergence Rate for TD(0) with Linear Function Approximation, Universal …
Learning of Robot Safety Policies via Adversarial Synthetic Scenarios
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Traject…
GLASS: GRPO-Trained LoRA for Acoustic Style Steering in Zero-Shot Text-to-Speech
QCFuse: Query-Aware Cache Fusion via Compressed View for Efficient RAG Serving
Entropy-Based Evaluation of AI Agents: A Lightweight Framework for Measuring Behavioral P…
Analysis of the Neglect-Zero Effect in Large Language Models
Inverse Design of Realizable Metasurface based Absorbers using Improved Conditioning and …
Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large L…
Can LLMs Be Constrained to the Past? Improving Knowledge Cutoff through Recall-Based Prom…
Beyond Absolute Scores: Relative Edit-induced Difference for Generalizable Image Aestheti…
Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents
MARDoc: A Memory-Aware Refinement Agent Framework for Multimodal Long Document QA
AdaPLD: Adaptive Retrieval and Reuse for Efficient Model-Free Speculative Decoding
Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models
Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solvi…
Seeing Time: Benchmarking Chronological Reasoning and Shortcut Biases in Vision-Language …
PerceptUI: LLM Agents as Human-Aligned Synthetic Users for UI/UX Evaluation
Benchmarking Counterfactual Prediction in Epidemic Time Series with Time-Varying Interven…
Data Flow Control: Data Safety Policies for AI Agents
Lost in Fog: Sensor Perturbations Expose Reasoning Fragility in Driving VLAs