ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Toward Better Assessment of LLMs' Performance in Clinical Error Detection
Position: AI Lock-In Is in Progress, and We Must Be Prepared
Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
Agentic-SQL Revisited: Autonomy-Based Taxonomy and Empirical Benchmark Analysis for LLM T…
Characterising cardiac tissue properties with graph neural networks
Position: Certified Correctness in Neural Constraint Reasoning Requires Symbolic Integrat…
pico-type: A 1.5M-Parameter Byte-Level Multi-Head Content Classifier
Defake-o3: From Speculative Rationales to Verifiable Evidence for Explainable AIGI Detect…
Protein Structure Prediction: From Evolutionary Constraints to Generative Modeling
Discovering High-Quality Chess Puzzles with Offline Reinforcement Learning
BaT: Towards Self-Evolving Medical Research Agent with Stage Rubrics
Do LLMs Know What to Ask and When? Evaluating Multi-Turn Information Seeking
FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment
Listen, Reason, and Segment: Aligning LALMs with Editorial Judgment for Media Chapterizat…
Hierarchical Adaptive Feature Refinement Network for VHR Remote Sensing Image Segmentation
Large Language Models Show Metacognitive Sensitivity in Medical Reasoning
From Errors to Proofs: Minimal-Core-Guided Repair for Neuro-Symbolic Constraint Solving
Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated …
Grounding Healthcare LLMs in a Causal Knowledge Graph: Framework, Metrics, and a Cardiova…
Robo-Dopamine 2.0: History-Conditioned and OOD-Aware Process Reward Modeling for Robotic …
Platform Adaptation Under Governance Interventions: Actor Best-Response Modeling and an E…
Prior Audit-Repair Context Shifts LLM Verifier Thresholds Toward Leniency
Auditing an AI-Generated Mathematical Proof: A Correction to a Greedy Conditioning Lemma …
Governance at the Boundary: How Agent Decomposition Degrades Policy Compliance
CardiacMamba: Fair and Robust RGB-RF Fusion for Remote Heart Rate Estimation via State Sp…
Incoherent by Design? On the Moral Self-Consistency of LLMs
A concentration result for multilayer feedforward neural networks
Semantic Uncertainty-Guided Orchestration in Hierarchical Multi-Agent Systems
Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynami…