More Than Mimicking Reviewers: Evaluating LLMs for Pre-Submission Peer Review
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Exposing Weaknesses in Emotion Recognition in Conversations
GIFT: Reconciling Post-Training Objectives via Variational Finite-Temperature Gibbs Initi…
Distilling Vision-Language Models for On-Device Fire Understanding
Agentic Pressure: The Endogenous Entropy of Reliable Autonomy
Modus Tollens and Counterfactuals and Counterfactual Reasoning Based on Three Types of Ne…
Quantile-Led Feature Extraction for Multi-Horizon Predictive Maintenance in Industrial Ma…
The Normalization of Deviance in AI Development
From Monolithic Blending to Agentic Orchestration: Dynamic Response for Conversational As…
Generator-Independent Runtime Assurance under Partial Observation
LEBGen: An LLM-Enhanced Bayesian Network Framework for Few-Shot Travel Survey Data Genera…
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record f…
The convergent laboratory: when AI reasoning, autonomous experiments, high performance an…
Recovering Temporal and Geographic Signals from Language Model Embeddings
CUSP: Decomposable Collective Uncertainty for Multi-Agent Multimodal Reasoning
CircuTutor: Transforming Static Circuit Problems into Intelligent and Dynamic Tutoring
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models
Towards Embodied Air-Ground Cooperative Object Search: Benchmark, Dataset and Agentic Met…
From Where to How: Continuous 4D Interaction Forecasting from Egocentric Video
SynthRCT: Scalable Conditional Deformation Synthesis for Synthetic Repeat CT Generation
Neptune: An AI model for Global Ocean Subseasonal Prediction
Deep belief networks are exact
Leveraging Cardiac Imaging to Improve ECG-Based Detection of Chagas Disease in Resource-C…
EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent
Graph-Based Personalized Memory for LLM Agents: Representation, Evolution, Retrieval, and…
Reasoning-Aware Compression: Identifying and Protecting Vulnerable Reasoning Circuits for…
SciLitBench: Benchmark and Design Principles for LLM-Powered Systematic Literature Reviews
The Failure Happens Before the Drift: The Social Dynamics of Values in LLM Agent Societies
Compiling VGDL into Causal Models
RAPID: Reliability-Aware Pair Importance Distillation