Recursive Criticality of AI Self-Improvement
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Reinforcement Learning Enhanced LLM Agents for Complex Vehicle Routing Problems
Flawed in Nature, Perfect through Evolution
Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline …
What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal
Why Fine-Tuning Encourages Hallucinations and How to Fix It
Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Eart…
Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis
Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs
Instella-MoE Technical Report
VectorGym: A Multi-Task Benchmark for SVG Code Generation, Sketching and Editing
REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models
Confess What You Know: Forget-Set Misalignment with Model Knowledge in LLM Unlearning
Training-Free Refinement of Flow Matching with Divergence-based Sampling
Designing Proactive Thought Partners for Writing
When the Algorithm Becomes the Brand Crisis: A Sociotechnical Theory of Distributed Respo…
Different representation learning objectives recover distinct latent structures from the …
Wave Function Backpropagation with Explicit Temporal-Interval Dynamics
Feedback-Assisted Trust Propagation over Document Relation Graphs for Retrieval-Augmented…
SCALE:Scalable Conditional Atlas-Level Endpoint transport for virtual cell perturbation p…
APEX-EM: Non-Parametric Online Learning for Autonomous Agents via Structured Procedural-E…
Consistency Without Alignment: Item-Sensitive Language Models Indistinguishable From Rand…
Channel-Adaptive Edge AI: Maximizing Inference Throughput by Adapting Computational Compl…
MMAI Gym for Science: Training Liquid Foundation Models for Drug Discovery
Reconstruct! Don't Encode: Self-Supervised Representation Reconstruction Loss for High-In…
FloydNet: A Learning Paradigm for Global Relational Reasoning
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
GeoPAR: Large-Scale Multi-Agent Combinatorial Optimization with Geometry-Guided Parallel …
EGT-KG: Evidence-Grounded Typed KG Retrieval for Practical Scientific QA with Small Langu…