Speaker-Disentangled Chunk-Wise Regression for Syllabic Tokenization
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Beyond Multilingual Averages: MTEB-PT, a Benchmark for Portuguese Sentence Encoders
Which Algorithm Specification Formats Help Language Models Implement Machine Learning Alg…
The Role of Prompt Language and Translation-Theory-Driven Prompts in Large Language Model…
DynaVieW: Schema-Guided World Modeling for Understanding Hierarchical Visual Dynamics
UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning
Generative wave propagator
AnchorVLA: Bridging Discrete Decisions and Continuous Trajectories for Vision-Language-Ac…
Scalable Maximal Frequent Episode Mining with Desbordante
Robustness Verification of an Autonomous Underwater Vehicle-based Plankton Classifier
Graph Neural Networks are Heuristics
Two Black Boxes, One Solver: Encoder Probing and Decoder Attribution for Neural Multi-Att…
A Bayesian Framework for Evaluating Scenario Compatibility in Generative Population Synth…
SoftVTBench: A Safety-Aware Visuo-Tactile Benchmark for Physically Constrained Robotic Ma…
Efficient Perception in Automotive Detection and Tracking Using Neuromorphic Computing
Joint Velocity Slope Diffusion Prior for Structurally Constrained Velocity Model Building
Detecting Hallucinations in Retrieval-Augmented Generation through Grounding-Aware Sensit…
Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environ…
CritiqueDriveVLM: From Verifier-Guided Reinforcement Learning to Latent Thought Distillat…
Failures and Successes to Learn a Core Conceptual Distinction from the Statistics of Lang…
Lyapunov-Guided Training for Hardware-Safe Neural Networks Under Fixed-Point Arithmetic
Mask2Real-WM: Segmentation Masks as a Sim-to-Real Bridge for Controllable Dexterous World…
EEG-SpikeAgent: Agentic Closed-Loop Program Synthesis for Automated EEG Spike Detection
SeeMe: Mitigating Hallucinations in Large Vision-Language Models through Effective Visual…
Mask-based Predictive Representations for Reinforcement Learning
RUFNet: Query-Guided Support Mask Refinement and Uncertainty Fusion based on Hybrid Mamba…
A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent P…
Simple-to-Complex Structured Demonstrations for Vision-Language-Action Learning
Builder, Defender, Breaker: The Case Against Removing the Human from the AI-Driven Securi…
Do All Visual Tokens Matter Equally? Object-Evidence Preserving Token Merging for Vision-…