LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platfo…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
FreKoo++: Learning Continuous Spectral Dynamics for Temporal Domain Generalization
Practical Principles for AI Cost and Compute Accounting
EG-ARSA: An Expert-Grounded Open Model for Visual Road Safety Auditing in Low-Resource Se…
SRMT: Shared Memory for Multi-agent Lifelong Pathfinding
Minimal Local Simulation Foundations for LLM- and VLM-Driven Agents in 2D and 3D Environm…
SkillAlchemy: Open-World Agent Skill Creation
FigmaTrace: Capturing Creative Nuances in Human Figma Design Workflows
Alignment midtraining for animals
The Emergence of Relevance Through Axiomatic Attention Patterns During LoRA Fine-Tuning
Small Reasoning Models are Instruction Followers in Function Calling
One-Step Evolution for Long-Time Extrapolation: An Error-Bound-Informed and Prior-Guided …
ChebBooster: A Training-Free Approach for Efficient Diffusion Transformer Inference via C…
PsychJail: Exploring Psychological Jailbreaks via Multi-Turn Persuasion of LLM Policies
CDEG: Learning Decision-Critical Evidence for Long-Horizon Diagnostic Agents
Proxy reliance in large language model decisions is uncalibrated to predictive evidence
RoboShape: Information-Theoretic Point Cloud Representations for Privacy-Aware Robot Perc…
A-CPES: A Reference Framework for Agentic AI in Cyber-Physical Energy Systems
How AI Assistance Affects Human Skill Development: A Study of Learning with Logic Puzzles
Concepts for Securing Agentic AI Coding and the Terok Environment
ProBel: Propaganda Detection with Techniques, Spans, and Explanations
GameXpert-Bench: How Far Are Coding Agents from Expert Game Development?
Towards a resource for multilingual lexicons: an MT assisted and human-in-the-loop multil…
Adaptive Item-based Collaborative Structures via Noise Rescheduling in Diffusion for Gene…
GSAR: Goal-State-Anchor Rewards for Mobile GUI Agents with Self-Evolving Data Synthesis
Let the Bullets Fly: Multimodal Fake News Detection with Temporal-Aligned Generative Danm…
What LLMs explain is not what they believe: Evaluating explanation sufficiency under mode…
AdaR: A Framework for Equipping LLMs with Adaptive Reasoning
WAM-OPD: On-Policy Distillation for World Action Models
GenCoord: Skill-Path Commitments under Private Information