NumLeak: Public Numeric Benchmarks as Latent Labels in Foundation Models
Explorar
Noticias de IA
29670 elementos — filtrados, clasificados y sin duplicados
Physically Viable World Models: A Case for Query-Conditioned Embodied AI
Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Au…
Structure-Induced Information for Rerooting Levin Tree Search
SLAT: Segment-Level Adaptive Trimming for Efficient CoT Reasoning
Distilling LLM Feedback for Lean Theorem Proving
Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Ag…
Diagnosing Failure Modes of Shared-State Collaboration in Resource-Constrained Visual Age…
G-STAR: End-to-End Global Speaker-Tracking Attributed Recognition
Rank-Factorized Implicit Neural Bias: Scaling Super-Resolution Transformer with FlashAtte…
Weight Decay Improves Language Model Plasticity
CodeGolf Bench: A Multi-Language Benchmark for Evaluating Concise Code Generation Capabil…
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
MAECO-Lite: Modular Ontology for Dynamic Malware Analysis
SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforc…
Toxic HallucinAItions: Perturbing Prompts and Tracing LLM Circuits
AMix-2: Establishing Protein as a Native Modality in Large Language Models
Learning Cardiac Latent Representations in Vectorcardiogram Space
Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning
Envisioning Beyond the Few: Disentangled Semantics and Primitives for Few-Shot Atypical L…
The Terminal Representation in Reinforcement Learning
Spectral Collapse Drives Loss of Plasticity in Deep Continual Learning
SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medica…
MedCoG: Maximizing LLM Inference Density in Medical Reasoning via Meta-Cognitive Regulati…
Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phr…
A Kinetic Energy Perspective of Flow Matching
FreeTimeGS++: Secrets of Dynamic Gaussian Splatting and Their Principles
Towards Atoms of Large Language Models
Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models