Do Modules Stay in Their Lane? Role Drift in Compound LLM Systems
Explorar
Noticias de IA
22114 elementos — filtrados, clasificados y sin duplicados
Defining AI-Native Systems: Autonomy as Revision Authority
Persistent Computational State: A Session-Centric Runtime for Generative World Models
HiKV: Hierarchical Importance-Aware KV Cache with Hardware Acceleration for LLM Decoding
Automatic Stability and Recovery for Neural Network Training
Hyperball May Not Be a Free Lunch
Unboxing Diffusion Models for the Arts: Interactive Model Bending and Practice-Based Expl…
What AI Red-Team Evaluations Can and Cannot Prove
From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for…
QLPO: Quadrant-weighted Sampling for Length-aware Policy Optimization
Co-design of LLM-based preference agents: participation may drive overtrust
DAGForge: Auditable Causal DAG Authoring with Biomedical Literature
When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual…
Multi-Agent System-driven Digital Twins for predictive maintenance: architectures, techno…
TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized…
Semiotic logical hexagon theory for LLM logical reasoning
Learning as Reasoning Unfolds: Progressive Rollout Allocation for Efficient Reinforcement…
Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reaso…
A Self-Calibrating Agentic AI Framework for Autonomous Edge Resource Allocation
Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for I…
CARDIAG: A Dense Segment Classification Benchmark of Deep Learning Architectures for Coro…
Learning on the Job: Continual Learning from Deployment Feedback for Frozen-Weights Agents
Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforce…
Benchmarking Text-to-SQL under Role-Based Access Control
AI4PLE: A Methodology for Integrating AI into Product Line Engineering
Re-FORC: Adaptive Reward Prediction for Efficient Chain-of-Thought Reasoning
DeepFeature: LLM-Empowered Context-aware Feature Generation for Wearable Biosignals
A Roadmap to Impactful Pluralistic Alignment Research
Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoenco…
From Mind to Machine: The Rise of Manus AI as a Fully Autonomous Digital Agent