RIACT: A Responsible AI System for Personalized Study Habit Tracking and Early Burnout Si…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Is Next-Chunk Reasoning RL Really Better than SFT? Revisiting Training Strategies under n…
CONTRAMEM: Learning Self-Evolving Procedural Memory from Contrasting Multi-Model Trajecto…
Towards Comprehensive Basketball Understanding
Procedural Knowledge Extraction from Industrial Troubleshooting Guides Using Vision Langu…
Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
Spending Scarce Confirmatory PET Measurements: Target-Aligned Validation in A4/LEARN
CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-…
ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Eva…
FlowExtract: Procedural Knowledge Extraction from Maintenance Flowcharts
On the Role of Citations in Preference Data
Language Chain in Alignment: Cross-Lingual Ranking Preference Optimization
Decision-Support and Modeling with Large Language Models for Geothermal Well Arrays
When Does AI for PDEs Yield Scientific Evidence?
More Accurate or More Efficient? Evaluating Locally Deployed Compact Open-Weight Language…
Sycophants in the Courtroom: Are LLMs Fragile to Juridical Authority and Evolving Legal S…
Evaluating Multimodal Narrative Understanding of Popular Hollywood Films
Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning
SchemaRouter: Field-Aware Tool Routing for Efficient Heterogeneous Agentic RAG
KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Effici…
Performance of a domain-specific large language model in answering patient questions in p…
ST-EVO: Towards Generative Spatio-Temporal Evolution of Multi-Agent Communication Topolog…
MegaMem: A Retrieval Solution for Ultra-Large Context Windows
StrategyBench: Evaluating Explicit Strategy Induction in Large Language Models
SAFE-G: Structure-aware Faithful Evidence-guided Generation for Knowledge-based Visual Qu…
Radial Compensation: The Inverse Base-Distribution Problem for Chart-Based Generative Mod…
From Recognition to Reasoning: Advancing Multimodal Harmful Meme Detection via Chain-of-T…
LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platfo…
LLM4LLM: Bridging Kernel Benchmarks and Real Deployment via Closed-Loop Agentic Optimizat…
EDGE: Experience-Distillation for Guided Exploration in Agentic Reinforcement Learning