FairFund-Bench: Evaluating Distributive Bias in LLM Resource Allocation
Explorar
Noticias de IA
30304 elementos — filtrados, clasificados y sin duplicados
Gated Q-learning: Add Off-Policy Bias to Taste
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feed…
On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness
Scaling Scientific Discovery Environments for Turn-Level Agentic RL
Identifying Informative Environments for Cognition Parameter Inference via Bayesian Exper…
Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures
EarlyDx: An Admission-Anchored Benchmark for Open-Ended Generation of Evidence-Supported …
SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Ac…
Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO
HenTwin: A Multimodal Digital Twin Framework for Longitudinal Biological State Monitoring…
COntExt: Towards Context-Aware Ontology Extension from Operational Metrics
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Rememb…
Beyond Retrieval: Analytic Memory for Multimodal Agents
Beyond Component Testing: Validating Agentic AI Systems
Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL
MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft
CAGE: Certified Authorization under Typed-Return Uncertainty for Tool-Using Agents
Adaptive Policy Backbone via Shared Network
AIvilization v0: Toward Large-Scale Artificial Social Simulation with a Unified Agent Arc…
A Generalized-Bayes Perspective on Counterfactual Explanations: Posterior-Based Decision-…
On a joint simultaneous learning of relevant feature subsets and subspaces in regression-…
LabEvolver: Training-Free Experience Evolution for Safe and Grounded Wet-Lab Agents
Design Concept: Scaffolding Geopolitical Reflection Among Tech Workers
TORUS: A Test of Rendering-Understanding Self-Coherence for Unified Audio Models
Agreement Is Not Quality: Blind Expert Verification of Human and LLM Qualitative Coding W…
DiffAttack: Evasion Attacks Against Face Recognition via Latent Diffusion Models
Retrieval-Driven Training-Free AI-Generated Video Attribution