Closing the Feedback Loop: From Experience Extraction to Insight Governance in Verbal Rei…
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
DeepInsight: A Unified Evaluation Infrastructure Across the Physical AI Stack
SEAGym: An Evaluation Environment for Self-Evolving LLM Agents
StepGuard: Guarding Web Navigation via Single-Step Calibration
FinAcumen: Financial Multimodal Reasoning via Self-Evolving Experience Memory Harness
Beyond Parallel Sampling: Diverse Query Initialization for Agentic Search
ReAge3D: Re-Aging 3D Faces with View Consistency
Descriptor: Certus Caliber Classification Gunshot Dataset (C3GD)
Embedded Machine Learning for Microcontroller-Class Edge Devices: Data, Feature, Evaluati…
ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model
Guidelines for the Annotation and Visualization of Legal Argumentation Structures in Chin…
Nothing from Something: Can a Language Model Discover 0?
Skill-Constrained Model Predictive Control for Resilient Manufacturing Supply Chains
MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guide…
SkillChain-Gym: A Benchmark for Reskilling-Aware Production-Inventory Control under Disru…
Co-PLNet: A Collaborative Point-Line Network for Prompt-Guided Wireframe Parsing
WallZero: Mastering the Game of WallGo with Strategic Analysis
DecoSearch: Complexity-Aware Routing and Plan-Level Repair for Text-to-SQL
SPATIA: Multimodal Generation and Prediction of Spatial Cell Phenotypes
Learning Cardiac Electrophysiology Digital Twins Through Agentic Discovery of Hybrid Stru…
WEQA: Wearable hEalth Question Answering with Query-Adaptive Agentic Reasoning
Your AI Travel Agent Would Book You a Bullfight: An Agentic Benchmark for Implicit Animal…
Knowledge Reutilization in Meta-Reinforcement Learning
STAR: SpatioTemporal Adaptive Reward Allocation for Text-to-Image RL Post-Training
Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty
DiagFlowBench: Evaluating How Language Models Handle Off-Procedure Inputs in Grounded Dia…
Learn to Quantify Social Interaction with Constraints for Pedestrian Walking
Rethinking Cross-Layer Information Routing in Diffusion Transformers
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learnin…