New Evidence, Same Choice: Testing Physical Experiment Selection in Vision Language Models
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
A Mathematical Theory of Pragmatic Information
BenchShield: Formal Model-Backed Instrumentation for Reward Integrity in LLM-Agent Evalua…
Topological Necessities: Mechanism-Invariant Strategic Subgoals for Cross-Embodiment Goal…
EGGROLL, Unrolled: Understanding and Improving Low-Rank Evolution Strategies at Scale
ReactHuman: A Physics-Grounded Benchmark for Human-Like Reactive Decision-Making in Embod…
Robust Multimodal Sentiment Analysis with Incomplete Modalities via Semantic-aware Comple…
Generative Marketing Mix Modeling: A Causal Inference Framework Linking GEO and GEM to Bu…
Tapes Together Strong: The Co-evolution of Computation and Cooperation
Multilingual in Name Only? Cultural and Linguistic Weaknesses of LLMs in Urdu
When Synthetic Data Hurts: On Catastrophic Forgetting in Skill Retrieval for LLM Agents
What a Random Draw from the MCP Registry Contains, and What Tool-Use Benchmarks Contain I…
CARTS: Contextual Autoregressive Rank Transcoding Steganography for Full-Capacity Keyed T…
Less can be More: What Aspects of Speech Drive End-of-Turn Detection
MindTopo: Can Foundation Models Reason in Topological Space?
Can Edge-Deployable Vision-Language Models Identify Species?
When Agents Disagree: Bayesian Backward Reasoning as a Label-Free Anchor for Multi-Agent …
Explainability Assistant: A Conversational XAI Interface for Interpreting Energy Consumpt…
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization
Role differentiation as ignition of a collective information engine: Structuration in Age…
Beyond Verified Answers: Solver-Informed Self-Distillation for Bootstrapping Operations R…
Autonomy, Social Norms, and Alignment: Towards a Developmental Framework for Autonomous A…
Lightweight LiDAR-Based Cone Detection Framework Using Random Forest for Formula Student …
Enabling Knowledge Graph Understanding at Scale with the EXplore Your Graphs ENgine (EXYG…
Prompt Revision as a Source of Cultural Bias in Text-to-Image Systems
From Document Silos to Process Intelligence: A Multi-Layer Knowledge Graph for CMC Proces…
The Truth Was Never Gone: Perfect Aliasing in Compliant-Context Truth Probes
ActMap: Single-Pass Uncertainty Quantification from Generation-Time Activation Maps
MAPLE: Memory-Augmented Planning with Language and Evolution
A machine-checked proof of the Dong-Yang classification of optimal (n,4) binary codes for…