Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent C…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Multimodal Drivers' Emotion Recognition and Safety-Oriented Intervention for Intelligent …
Mitigating Scoring Bias in LLM-as-a-Judge via Random Number Generation
SkillProx: Self-Evolving Agent Skills via Proximal Textual Gradient Descent
Blast Radius
PsychoAgent: An Affect-Sensitive Cognitive Architecture for Conflict-Aware Memory in LLM …
Post-Grokking Collapse at the Representation-Readout Interface in Muon-Trained Transforme…
TEPA: Revoking Stale Memories for Conflict-Robust Language Agents
A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs Whil…
ResidencyRL: Reinforcement Learning in Simulated Clinical Environments
An End-to-End Agent Auditing Engine
Winning by Peeking: Unenforced Budgets and Test-Set Selection Inflate Short-Budget AutoML…
GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks
FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval ov…
Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education
WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
From probability to causality in probabilistic logic programming
Authoring and Management of Transparent Research Integrity Assessments of Randomised Clin…
SetEasy: A Multi-Modal Classroom Engagement Assessment and Seating Optimization Framework
NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zer…
A MARL Centered Reference Architecture for Large Language Model Augmentation in Smart Man…
DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training
How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent R…
Transformers Struggle to Use Their Emergent World Models: Revisiting the Tower of Hanoi, …
DocMemo: Dynamic Evidence Discovery via Probabilistic Memory-Guided Retrieval for Multi-M…
BONSAI: Evolvability-Guided Tree Search over Skills
Unsupervised Adaptation of PDE Foundation Models
Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accur…
ReQuant: Fixed-Grid Discrete Refinement for Post-Training Quantization
Finding Usable Weight Mechanisms with Tiled SVD