ImProver 2: Iteratively Self-Improving LMs for Neurosymbolic Proof Optimization
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Any2Any: Efficient Cross-Embodiment Transfer for Humanoid Whole-Body Tracking
Redrawing the AI Map: A Theory of Accountability Boundaries in Agentic Ecosystems
Model Spec Midtraining: Improving How Alignment Training Generalizes
GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models
Socially fluent AI decouples conversational signals from source identity in online intera…
Design and Report Benchmarks for Knowledge Work
Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preferen…
STM3: Mixture of Multiscale Mamba for Long-Term Spatio-Temporal Time-Series Prediction
Fine-Tuning Causal LLMs for Text Classification: Embedding-Based vs. Instruction-Based Ap…
BarrierSteer: LLM Safety via Learning Barrier Steering
HTMuon: Improving Muon via Heavy-Tailed Spectral Correction
MUSEKG: A Knowledge Graph Over Museum Collections
V-VLAPS: Value-Guided Planning for Vision-Language-Action Models
An Interpretable Closed-Loop Intelligent Tutoring System for Multimodal Affective Feedbac…
Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling
Learning Through Noise: Why Subliminal Learning Works and When It Fails
Mediative Fuzzy Logic: From Type-1 Foundations to Type-2, Type-3 and Quantum Extensions
Empowering 9-1-1 Calltaking Training with Generative AI: Experiences and Lessons Learned
Convergence Without Understanding: When Language Models Agree on Representations but Disa…
Representational Alignment with Chemical Induced Fit for Molecular Relational Learning
Evaluating Large Language Models in a Complex Hidden Role Game
Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robus…
Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA…
S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination
It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amp…
MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structura…
PGT: Procedurally Generated Tasks for improving visual grounding in MLLMs