Measuring Intelligence Beyond Human Scale
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
SOMtime the World Ain$'$t Fair: Violating Fairness Using Self-Organizing Maps
Large Behavior Model: A Promptable Digital Twin of the Retail Customer
The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agent…
Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics
Reliable and Developer-Aligned Evaluation of Agents for Software Engineering
AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning
A Continual Learning Framework for Adaptive Control of Modular Soft Robots
A Multi-Analyst LLM Pipeline for Auditable Rule Discovery Across 68 Public Physiological …
AirPASS: Over-the-Air Federated Learning via Pinching Antenna Systems
Learning social norms enhances compatibility in dynamic human-AI coordination
Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulatio…
SPEAR: A Simulator for Photorealistic Embodied AI Research
What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study
Enhancing deep learning models for time series classification via knowledge distillation
ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits
Can We Really Learn One Representation to Optimize All Rewards?
From Agentic to Autogenic Network Management for AI-Native 6G and Beyond: A Standards Per…
LipSSD: Lipschitz-Constrained Single-Shot Detection for Adversarially Robust Object Detec…
CompDiff: Hierarchical Compositional Diffusion for Fair and Zero-Shot Intersectional Medi…
Effective Strategies for Asynchronous Software Engineering Agents
HiDVFS: Hierarchical Multi-Agent DVFS for Real-Time OpenMP DAG Workloads
SmartHomeSecure: Automated Detection and Repair of Smart Home Configuration Errors Using …
QCNN with Rough Path Signature Kernels
Exploration of Fast-Slow Latent Recurrence for Train-Short, Test-Long Generalization
L-GTA: Latent Generative Modeling for Time Series Augmentation
Deep Learning Method for Stationary Distribution of Reflected Brownian Motion
Recruiters Shift Focus to Specialized AI Jobs to Stay Relevant
Mixture of Enhanced-View Experts for Multi-Query Vehicle ReID and A Large-Scale Benchmark