Enhancing Decision-Making with Large Language Models through Multi-Agent Fictitious Play
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Optimal scenario design for climate emulation
Does VLA Even Know the Basics? Measuring Commonsense and World Knowledge Retention in Vis…
Correct Yourself, Keep My Trust: How Self-Correction and Social Connection Shape Credibil…
Transformer Geometry Observatory TGO-I: Spectral Geometry Observatory
Report: AI can perpetuate anti-LGBTQ hate
A Human-in-the-Loop Bayesian Optimization Framework for Constraint-Aware Bioprocess Devel…
Mechanism-Guided Selective Unlearning for RLVR-Induced Reasoning
Machine Unlearning for the XGBoost Model with Network Intrusion Datasets
GUMP-Net: An interpretable model-data-driven intelligent algorithm for multi-class pelvic…
Generalised Eigenvalue Geometry of Semantic Adversarial Attacks
MolmoMotion: Language-guided 3D motion forecasting
New research shows how AMIE, our medical AI, could help manage health conditions.
IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language …
AdsMind: A Physics-Grounded Multi-Agent System for Self-Correcting Discovery of Adsorptio…
On Local Population-Risk Certificates
Giskard : Byzantine Robust and Confidential Aggregation for Large-Scale Decentralized Lea…
Towards an Agent-First Web: Redesigning the Web for AI Agents
JourneyFormer: Encoding Airbnb Guest Journey with Sequence Modeling
Smoothness-Based Derandomization of PAC-Bayes Bounds
Structure Over Nonlinearity: Explicit Interaction Architectures for Dynamical Learning
Sensor Configuration Matters: A Systematic Evaluation of Multimodal SLAM on Quadruped Rob…
Which Sections of a Research Paper Best Reveal Its Research Methods? Evidence from Librar…
G-IdiomAlign: A Gloss-Pivoted Benchmark for Cross-Lingual Idiom Alignment
Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Rea…
FOSC-X: An Extended Framework for Optimal Local Cuts and Non-Horizontal Cluster Selection…
A Controlled Benchmark of Quantum-Latent GAN Augmentation for Brain MRI
Be Your Own Teacher: Steering Protein Language Models via Unsupervised Reward Optimization
SenFlow: Inter-Sentence Flow Modeling for AI-Generated Text Detection in Hybrid Documents
SciRisk-Bench: A Risk-Dimension-Aware Benchmark for AI4Science Safety