The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
From Document Silos to Process Intelligence: A Multi-Layer Knowledge Graph for CMC Proces…
Published Unlearning Numbers Move Per Checkpoint, and Not Because the Removed Data Surviv…
ActMap: Single-Pass Uncertainty Quantification from Generation-Time Activation Maps
AcFlow: Controlling Text-to-Image Diffusion Transformers via Learned Conditional Activati…
Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYG…
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization
Recognizing Is Not Reversing: A Controlled Inversion Test of Fact-Preserving News Framing
When Agents Disagree: Bayesian Backward Reasoning as a Label-Free Anchor for Multi-Agent …
LLM-Ideoplasticity: Measuring Ideological Plasticity in the Political Behavior of LLMs as…
LLMs as Post-hoc Auditors of Physiological Plausibility in Symbolic Regression: A Clinici…
Flexible and Interpretable Accent Distance Measurements
Mr.LHDR: A Benchmark for Multimodal Real-World Long-Horizon Deep Research Agents
Magenta: Closing the Loop Between Mathematical Reasoning and Lean Verification
RAMamba-Net: A Reliability-Aware and Mamba-Based Multimodal Fusion Network for Auditory A…
Memory Compression for High-Fanout Agent Sandboxes
AI Exposure and AI Resilience: A Two-Dimensional Assessment Framework for Software and So…
Generating a Consistent Enterprise: Synthesis and Reference-Free Evaluation of Multi-Syst…
Buyer Artificial Intelligence-Enabled Environmental Governance and Supplier Environmental…
Bio-inspired Learning and Decision-Making with Probabilistic In-Memory Computing Hardware…
RouteRepair: Instance-Level Failure Diagnosis and Targeted Repair in LLM-Based Automated …
From Queries to Narratives: Cultural Heritage Data Stories for Knowledge Graph Exploratio…
Sci-MMR: Benchmarking Multi-Step Evidence-Grounded Scientific Reasoning in Multimodal Age…
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
2AM: Grounding Agent-Side Memory as Guidance for Steerable Action Models in Long-Horizon …
Story Imprinting: AI Assistants Absorb Traits from Human Characters They Resemble
ZipCodec: Ultra-Low-Frame-Rate Streaming Speech Coding
When Does Text Inform? Benchmarking Information-Theoretic Metrics for Multimodal Time-Ser…
Calibration-Aware Uncertainty Cascades for Efficient Heterogeneous Model Collaboration
Debate-to-Skill: Capability-Bound Process Supervision for Industrial Query-to-Agent Annot…