Differentially Private Preference Data Synthesis for Large Language Model Alignment
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforc…
MechVQA: Benchmarking and Enhancing Multimodal LLMs on Comprehensive Mechanical Drawing U…
OpenSTBench: Beyond Semantic Evaluation for Speech Translation
On the impact of retrieved content representations in RAG Pipelines
OrcaRouter: A Production-Oriented LLM Router with Hybrid Offline-Online Learning
GSAM: A Generalizable and Safe Robotic Framework for Articulated Object Manipulation
SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs
Chatterbox-Flash: Prior-Calibrated Block Diffusion for Streaming Zero-Shot TTS
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM U…
SWIM: Single-Instance Whole-Body Imitation for swiMming
Rank-Factorized Implicit Neural Bias: Scaling Super-Resolution Transformer with FlashAtte…
Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence
MAECO-Lite: Modular Ontology for Dynamic Malware Analysis
MIMO: Multilingual Information Retrieval via Monolingual Objectives
Simple Token-Efficient Vision-Language Model for Case-level Pathology Synoptic Report Gen…
Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reve…
Depth-Dependent Indirect Prompt Injection in Tool-Calling ReAct Agents: Injection Depth, …
Learning to Solve and Optimize by Evolving Code
Human-Alignment, Calibration, and Activation Patterns in Large Language Model Uncertainty
Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clin…
LARK: Learnability-Grounded Trajectory Selection for Efficient Reasoning Distillation
EUDAIMONIA: Evaluating Undesirable Dynamics in AI
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
Rationalize: Shared Semantic Reasoning for Human-AI Alignment
PInVerify: An Offline Embodied Benchmark for Active Instance Verification
Developing an AI-Powered UX Research Point of View for Digital Health in A Regulatory Con…
KnowledgeGain: Evaluating and Optimizing Science News Generation for Reader Learning
Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inp…