Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resourc…
Explorar
Noticias de IA
29629 elementos — filtrados, clasificados y sin duplicados
RoboBenchMart: Benchmarking Robots in Retail Environment
Understanding the Effects of Distractors on Reasoning Vision-Language Models
SpeedAug: Policy Acceleration via Tempo-Enriched Policy and RL Fine-Tuning
From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language…
LALE: Lightweight-Transformer Architecture for Land-Cover Estimation
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents
InFerActive: Interactive Tree-Based Exploration of LLM Sampling for Safety Evaluation
Why Do Time Series Models Need Long Context Windows?
Can Vision Models Truly Forget? Mirage: Representation-Level Certification of Visual Unle…
Co-Fusion4D: Spatio-temporal Collaborative Fusion for Robust 3D Object Detection
The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue
Physics-Guided Geometric Diffusion for Macro Placement Generation
RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation
Beyond AI as Assistants: Toward Autonomous Discovery in Cosmology
Control of a Twin Rotor using Twin Delayed Deep Deterministic Policy Gradient (TD3)
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
RISED: A Pre-Deployment Evaluation Framework for High-Stakes AI Decision-Support Systems,…
Multi-Rollout On-Policy Distillation via Peer Successes and Failures
OGLS-SD: On-Policy Self-Distillation with Outcome-Guided Logit Steering for LLM Reasoning
Normalization Equivariance for Arbitrary Backbones, with Application to Image Denoising
Do Joint Audio-Video Generation Models Understand Physics?
Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA
MGRegBench: A Novel Benchmark Dataset with Anatomical Landmarks for Mammography Image Reg…
PBT-Bench: Benchmarking AI Agents on Property-Based Testing
STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems
Possibilistic Predictive Uncertainty for Deep Learning
Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
FlowPlace: Flow Matching for Chip Placement