HybridCodeAuthorship: A Benchmark Dataset for Line-Level Code Authorship Detection
Explorar
Noticias de IA
22116 elementos — filtrados, clasificados y sin duplicados
Analyzing and Improving Fine-grained Preference Optimization in Medical LVLMs
Graph Reduction in Multirelational Networks: A Spreading-Oriented Reduction Benchmark
Brick: Spatial Capability Routing for the Mixture-of-Models (MoM) Paradigm
A Mathematical Theory of Value: a synthesis on goal-directed agency under resource constr…
Improving Crash Frequency Prediction from Simulated Traffic Conflicts Using Machine Learn…
Physics-Guided Spatiotemporal Learning for Coastal Wave Peak Period Estimation from Video
Speculative Rollback Correction for Quality-Diverse Web Agent Imitation
ReCal: Reward Calibration for RL-based LLM Routing
Multi-Modal Agents for Power Distribution Defect Detection: An Evaluation of Foundation M…
A Mathematical Forum Platform for Collaborative Problem Solving and Dataset Generation fo…
ToolSense: A Diagnostic Framework for Auditing Parametric Tool Knowledge in LLMs
ASTER: Latent Pseudo-Anomaly Generation for Unsupervised Time-Series Anomaly Detection
GeoDial: A Multimodal Conversational Tutoring Dataset for Geometry Problem-Solving with V…
AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and Reproducibili…
Prism: Cost-Efficient Multi-LLM Serving via GPU Memory Ballooning
WildIFEval: Instruction Following in the Wild
APCyc: Property-Informed Design of Cyclic Peptides via Automated Cyclization
SciR: A Controllable Benchmark for Scientific Reasoning in LLMs
Augmentation techniques for video surveillance in the visible and thermal spectral range
Rethinking RAG in Long Videos: What to Retrieve and How to Use It?
Automated reproducibility assessments in the social and behavioral sciences using large l…
Proprioceptive-visual correspondence enables self-other distinction in humanoid robots
Mapping AI Programs in the U.S: A Status Report from Early 2026 and an Analysis of AI Maj…
Will AI Agents Free Us From Meaningless Work? A Human-Centered Analysis
Generativism: Toward a Learning Theory for the Age of Generative Artificial Intelligence
Occupational Prompting Reveals Cultural Bias in Large Language Models
Reasoning as Pattern Matching: Shared Mechanisms in Human and LLM Everyday Reasoning
Quickest Detection of Hallucination Onset: Delay Bounds and Learned CUSUM Statistics
Representing Time Series as Structured Programs for LLM Reasoning