The Fragility of Strategic Thinking in Large Language Models
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Calibrated Generative AI as Meta-Reviewer: A Systemic Functional Linguistics Discourse An…
Information Geometry of Message Passing
Proteus: Incremental Memory Activation for Long-Context Sequence Modeling
mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA
Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and…
SymbolicLight V1: Spike-Gated Dual-Path Language Modeling at High Activation Sparsity
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
AutoSR: Automatic Symbolic Regression by Searching Research States
NICE: Scale-Stable Perturbations for Graph Neural Network Explanations via Noise Corrupti…
Counting Documents Is Not Counting Text: Unit Bias in Web-PDF Corpus Statistics
Matched Outcomes, Divergent Gaze: How Foveated MLLMs Search Compared to Humans
Toward AI-Friendly Cartography: Understanding How Color Design Influences Foundation Mode…
HyMem: Hierarchical Context Management for Long-Horizon Agents via Information Isolation
Adaptive Mixing of Policies from Searching and Policies from Learning
Large Models for Small Devices: Recent Advances and Empirical Analysis of Edge AI Deploym…
Decoupling Parcellation from Classification: Systematic Benchmark of Fast Brain Segmentat…
Walk Before You Run: The Importance of Data Exploration for Data Analysis Agents
Model Hypnosis: Strong control of AI via additive subliminal effects
Command-Space Counterfactual Explanations for Pareto-Conditioned Reinforcement Learning
Do Geometry-Aware Positional Encodings Help Transformers in Spatial Imperfect-Information…
PAS-QFL: Personalized Ansatz Selection for Quantum Federated Learning under Client Data H…
Diagnosing Dense Same-Class Attribute Misbinding in Large Vision-Language Models
UniDot: A Unified Network for Sequence Modeling and Feature Interaction in Large-scale Re…
Historical Backtesting for Scientific Question Discovery: A Protocol and Astronomy Pilot
CAPO: Constraint-Aware Prompt Optimization for LLM Agents
Behaviour Is an Incomplete Measure of Reasoning Development: Cross-surface pre-arrival ac…
The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perc…
When to Communicate: Belief Distributions and KL Divergence for Principled Gating in Mult…
From Doyle to AGM: A Survey and an Implementation Roadmap for Belief Change