Accelerating Skill Assessment in Chess: A Drift-Diffusion-Enhanced Elo Rating System
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Gove…
Refusal Lives Downstream of Persona in Chat Models
Life After Benchmark Saturation: A Case Study of CORE-Bench
IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in M…
SpaceRipple: Lightweight Semantic Delivery for Mission-Oriented LEO Earth Observation Sat…
\textsc{DiARC}: Distinguishing Positive and Negative Samples Helps Improving ARC-like Rea…
Revisiting the Platonic Representation Hypothesis: An Aristotelian View
Improved Bounds for Private and Robust Alignment
Patent Representation Learning via Self-supervision
Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Languag…
Byzantine-Robust Aggregation for Securing Decentralized Federated Learning
Wearable Device-Based Real-Time Monitoring of Physiological Signals: Evaluating Cognitive…
Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Lib…
Joint Reward Modeling: Internalizing Chain-of-Thought for Efficient Visual Reward Models
Beyond the Hard Budget: Sparsity Regularizers for More Interpretable Top-k Sparse Autoenc…
AI Healthcare Chatbots as Information Infrastructure: A Large-Scale Study of User-Reporte…
Rotary Position Encodings for Graphs
NuclearQAv2: A Structured Benchmark for Evaluating Domain-Science Competence in Large Lan…
[AINews] OpenAI reports median internal Codex output tokens grew 56x in Research, 32x in …
Training Observable Control Policies to Expose Agent State Through Actions
Don't Settle at the Mode! Mitigating Diversity Collapse in Pretrained Flow Models via Fea…
Autoregressive Boltzmann Generators
Retrofit, don’t rebuild: Agentic overlays for transforming legacy enterprise services
Hallucination in World Models is Predictable and Preventable
Multilingual Reasoning Cascades Need More Context
Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN
A Multi-Fidelity Convolutional Autoencoder-Transfer Learning Framework for Guided-Wave-Ba…
Designing Reward Signals for Portable Query Generation: A Case Study in Industrial Semant…
Recovering Governing Equations from Solution Data: Identifiability Bounds for Linear and …