STAIR: Effective Incident Response Using an End-to-End Agentic Planning Framework
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
MARA: Flow-Matching-Guided Multi-Agent Resource Allocation for Computational Resource Eff…
SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning
iLTM: Integrated Large Tabular Model
Automated Generation of Complexity-Validated Decision Scenarios Using Large Language Mode…
Monotonicity-Guided Bottom-Up Petri Net Discovery: The SPECpp Framework
Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounde…
PROSLEX: A Novel Dataset for Expert-Annotated Legal Statute Prediction for Indian Judicia…
Findings of the First Teaching Monster Challenge: A Benchmark of Pedagogical Content Know…
Imaginative Generative AI: Crossing the Entropy Wall into Worlds Beyond Imitation
Second Order Drifting Models
EasyBalance: Cross-Layer Load Balancing in Distributed MoE Inference
Thinking Hard, Not Smart: Reasoning Models Fail to Ration Test-Time Compute Across Questi…
Evidence-RL: Towards Evidence-intensive Visual Reasoning
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safe…
DialectS2S: End-to-End Speech Dialogue Modeling for Low-Resource Chinese Dialects
Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizati…
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching
DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task L…
STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in …
Population-Scalable Multi-Agent World Modeling
AquiLLM: An Architecture for Supporting Tacit Knowledge Capture in Research Groups
Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Pro…
Full-bandwidth transformer
Integrated Multimodal AI System for Retrieval-Augmented Reasoning, Object Sensing, and Da…
Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM
Three Necessary Principles for Self-Supervised Visual Representation Learning
Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution