Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal M…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
JEPA-DNA: Grounding Genomic Foundation Models through Joint-Embedding Predictive Architec…
SEA-Eval: A Benchmark for Evaluating Self-Evolving Agents Beyond Episodic Assessment
Why Your Deep Research Agent Fails? On Hallucination Evaluation in Full Research Trajecto…
AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent
Dynamic Dual-Granularity Skill Bank for Agentic RL
Generative structure search for efficient and diverse discovery of molecular and crystal …
Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Sp…
Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending
CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures
Efficient Benchmarking Is Just Feature Selection and Multiple Regression
Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork
Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning
TIGER: Text-Informed Generalized Enzyme-Reaction Retrieval
DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with…
Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qu…
Retrying vs Resampling in AI Control
Positivity in classical enumerative geometry: a case study in synchronized AI-assisted ma…
From Model Scaling to System Scaling: Scaling the Harness in Agentic AI
LETS Forecast: Learning Embedology for Time Series Forecasting
Constraint-Anchored Attribution: Feasibility-Certified Counterfactuals and Bonferroni-PAC…
Document Classification Pattern Recognition via Information Fusion: A Systematic Review o…
Agent-Facing Information Design in LLM Tool Registries
Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Pe…
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
On the Epistemic Uncertainty of Overparametrized Neural Networks
From Latent Space to Training Data: Explainable Specialization in Minimal MLPs
By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They…
Hide to Guide: Learning via Semantic Masking