Inverse Theory of Mind Modeling for Content Recommendation: From Web Browsing to Dynamic …
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs
On the Definition of Intelligence
VICBench: A Multi-Language Benchmark for Code Vulnerability Detection
EvoGraph-Mem: Failure-Aware Editable Graph Memory for Long-Term Language Agents
One Frozen Simulator Is Not Enough: Simulator Collapse in Multi-Agent RL
Learning from Online User Feedback for Shopping Agents
Physics-Informed Implicit Neural Representations for Improved Myocardial Perfusion MRI Qu…
InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discove…
ExRole: From Team Trajectories to Executable Roles in Multi-Agent Language Models
HyperANFIS: Enhancing Rule Representation and Interpretability in Adaptive Neuro-Fuzzy Sy…
CLAIM: Leading Open-domain Active Clarification of Large Language Models with Uncertainty…
GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Se…
Retry, Switch, or Abstain? Learning Strategy-Aware Tool-Use Policies via Controlled Error…
MaSRead: Content-Addressed Reading of Replicated Latent Stores
TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation
Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended …
Variable Selection in the Context of AI Fairness
Beyond Fixed Luminance: Towards Panchromatic and Orthochromatic Image Colorization
How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Mode…
Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Con…
Ethics Practices in AI Development: An Empirical Study Across Roles and Regions
Quantization-Aware Neuromorphic Architecture for Skin Lesion Classification on Resource-C…
Program Semantic Inequivalence Game with Large Language Models
DORA Explorer: Improving the Exploration Ability of LLMs Without Training
ENTLORE: A Graph-Grounded Benchmark for Latent Organizational Reasoning in Enterprise Que…
Small Data Explainer -- The impact of small data methods in everyday life
Post-Training with Policy Gradients: Optimality and the Base Model Barrier
Representation Finetuning for Continual Learning