ATLAS: Scaffold-Free Algorithm Synthesis by LLMs via Embedding-Guided Quality-Diversity S…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
ReasonCast: Agentic Demand Forecasting with Selective Semantic Reasoning
Who Leads Now? Token-Level Modality Arbitration for Chart-to-Code Generation
The Benchmark Trap: Structures of Power and Injustice in AI Evaluations
CardiacMamba: Fair and Robust RGB-RF Fusion for Remote Heart Rate Estimation via State Sp…
A Network-driven Framework for Public Event Forecasting via Dynamic Interaction Network E…
A Responsible Artificial Intelligence Framework for Groundwater Modeling
Bias-Corrected Ceilings of Emotion Predictability from Human Label Variation Based on Ins…
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language…
Reasoning-supported Robustness Validation of Automotive E/E Components
Position: Certified Correctness in Neural Constraint Reasoning Requires Symbolic Integrat…
Think Inside the Chunk: RegulaRAG for Regulation-Compliant Scenario Generation using LLMs…
Defake-o3: From Speculative Rationales to Verifiable Evidence for Explainable AIGI Detect…
Large Language Models Show Metacognitive Sensitivity in Medical Reasoning
Position: AI Lock-In Is in Progress, and We Must Be Prepared
Discovering High-Quality Chess Puzzles with Offline Reinforcement Learning
Do LLMs Know What to Ask and When? Evaluating Multi-Turn Information Seeking
Argumentation for Common Ground: Finding Zones of Possible Agreement between Individuals …
LongDocBench: Benchmarking TOC Hierarchy and Contextual Relationship Recovery in Long Doc…
From Errors to Proofs: Minimal-Core-Guided Repair for Neuro-Symbolic Constraint Solving
Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated …
Platform Adaptation Under Governance Interventions: Actor Best-Response Modeling and an E…
Rotation-Invariant Multi-IMU Activity Recognition under Independent Per-Location Orientat…
FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment
PolyComp: A Polycube-based Benchmark for Compositional 3D Spatial Reasoning in Multimodal…
Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning
JarvisBench: Always-on Intelligence Between Humans and Agents
Beyond Asking: A Pipeline for Personalized Game Generation that Reads Players from Behavi…
Process-Constituted Intelligence: A Shared Criterion for Humans and Machines
Position: AI Agents in Scientific Teams Should Be Studied as Human-Agent Systems