From Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
How Compliant is Sepsis Treatment? An Expert-Guided Neuro-symbolic Pipeline for Generatin…
Adaptive Stopping for Multi-Turn LLM Reasoning
Knowing When to Stop: Bayesian Optimal Stopping for LLM Evaluations
A Hybrid LLM-Based Framework for Automated Security Annotation Generation in Business Pro…
Shift Aware Transfer Learning with Adaptive Dual-Encoder Fusion for PM Forecasting in Dat…
Participatory Moral AI Is Not Neutral: The Invisible Hand of Developers
GBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack Detection
AgentRewind: Recoverable Execution for Long-Horizon LLM Agents
PACE-Bench: Benchmarking Physics Adaptation via Code Evolution in Dynamic Environments
Wyvern: An Agentic Framework for Generating Grounded Multimodal Reports
Reinforcement Learning-Based Production Scheduling in an Industry-Based Coating Scenario …
A Graph-Based Reinforcement Learning Framework for Structured Drift Diagnosis and Recover…
FreeBalance: Pre-Routing Online Moe Load Balancing via Residual Workload Prediction
A Systematic Comparison of Training Objectives for Out-of-Distribution Detection in Image…
Modular Cognitive Architecture Emerges in Large Language Models
Stable Miscalibration in Large Language Models: A Practical View of High-Confidence Errors
MobileMem: Learning from a Year of Mobile Experiences
Measuring Fairness in Large Audio Language Models via Semantic-Aware Bias Estimation
CFM-Bench: A Unified Multi-Domain, Multi-Task Benchmark for Channel Foundation Models
VISOR: Agentic Visual Retrieval-Augmented Generation via Iterative Search and Over-horizo…
Algorithm Design and Physician Liability
TimeSage-EV: A Live Benchmark for Agentic Time Series Analysis in Evolving Environments
Not All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM Systems
Evaluating Agentic Learning Harness Capabilities Without Labels via the Scaling Hypothesis
Structured Prediction for Scalable Spreadsheet Table Understanding: From Cell Types to Ta…
New policy ideas for the Intelligence Age
$R^3$-Bench: LLMs Struggle with Resource-Rational Reasoning under Shared Budgets
AdROD: HyperNetwork-based Adversarially Robust Object Detection for Autonomous Driving
Multi-scale Decomposed Convolution Refinement Network for Visible-Infrared Person Re-Iden…