Fundamental Limitation in Explaining AI
Browse
AI News
30934 items — filtered, classified, deduplicated
When Mean CE Fails: Median CE Can Better Track Language Model Quality
Agent-Facing Information Design in LLM Tool Registries
Document Classification Pattern Recognition via Information Fusion: A Systematic Review o…
LETS Forecast: Learning Embedology for Time Series Forecasting
From Model Scaling to System Scaling: Scaling the Harness in Agentic AI
Retrying vs Resampling in AI Control
Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qu…
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with…
DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations
TIGER: Text-Informed Generalized Enzyme-Reaction Retrieval
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
Efficient Benchmarking Is Just Feature Selection and Multiple Regression
Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending
Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Sp…
Generative structure search for efficient and diverse discovery of molecular and crystal …
Dynamic Dual-Granularity Skill Bank for Agentic RL
SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment
JEPA-DNA: Grounding Genomic Foundation Models through Joint-Embedding Predictive Architec…
Why Your Deep Research Agent Fails? On Hallucination Evaluation in Full Research Trajecto…
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures
Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning
Positivity in classical enumerative geometry: a case study in synchronized AI-assisted ma…
Constraint-Anchored Attribution: Feasibility-Certified Counterfactuals and Bonferroni-PAC…
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
On the Epistemic Uncertainty of Overparametrized Neural Networks
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They…
Hide to Guide: Learning via Semantic Masking
Remote sensing data imputation using deep learning for multispectral imagery