AI evaluation may bias perceptions: The importance of context in interpreting academic wr…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Self-Improvement Imitation with Biologically Guided Search for Protein Design Under Oracl…
The Necessity of a Unified Framework for LLM-Based Agent Evaluation
Hi-SAM: A Hierarchical Structure-Aware Multi-modal Framework for Large-Scale Recommendati…
SL-BiLEM: Structured Learnable Behavior-in-the-Loop Epidemic Modeling for Forecasting and…
Stabilizing Recurrent Dynamics for Test-Time Scalable Latent Reasoning in Looped Language…
RAGEAR: Retrieval-Augmented Graph-Enhanced Academic Recommender
Periodic Topological Deep Learning for Polymer Design and Discovery
The Sensation Modulating Network:Haltability as the architectural ground for object-direc…
Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations
GeoFaith: A Spatio-Temporal Dual View of Faithful Chain-of-Thought
GSM-SEM: Benchmark and Framework for Generating Semantically Variant Augmentations
ICICLE: Expanding Retrieval with In-Context Documents
Practical Anonymous Two-Party Gradient Boosting Decision Tree
Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation ac…
JuICE: A Benchmark for Evaluating LLM-Judge in Identifying Cultural Errors
Cast a Wider Net: Coordinated Pass@K Policy Optimization for Code Reasoning
Evaluating the Relevance of Uncertainty Estimators for LLM Hallucination
Trust Region Q Adjoint Matching
DEI: Diversity in Evolutionary Inference for Quality-Diversity Search
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentat…
Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
Governed Evolution of Agent Runtimes through Executable Operational Cognition
When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Det…
GENESIS: Harnessing AI Agents for Autonomous 6G RAN Synthesis, Research, and Testing
AI-Driven Contribution Evaluation and Conflict Resolution: A Framework & Design for Group…
PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic
Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic