World Feedback for Clinical Agents: Diagnosing RL in FHIR Environments
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Fast Multi-dimensional Refusal Subspaces via RFM-AGOP
Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignm…
ADMC: Attention-based Diffusion Model for Missing Modalities Feature Completion
Causal Explanations for Image Classifiers
Procedural Memory Distillation: Online Reflection for Self-Improving Language Models
Decomposer: Learning to Decompile Symbolic Music to Programs
CreativityNeuro: Steering Language Model Weights to Improve Divergent Thinking and Reduce…
Discrete Diffusion Language Models for Interactive Radiology Report Drafting
LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning
The Agentic Garden of Forking Paths
Janus: a Playground for User-Involved Agentic Permission Management
Revisiting Chain-of-Thought Reasoning under Limited Supervision: Semi-supervised Chain-of…
OPINE-World: Programmatic World Modeling with Ontology-error-Prioritized Interactive Expl…
Scaling Trends for Lie Detector Oversight in Preference Learning
Prompt Framing Distorts Count-Based Evaluation of LLM Error Detection: Evidence from Nume…
ExPerT: Personalizing LLM Responses to Users' Domain Expertise via Query-Wise Semantic an…
LLMs as Teaching Assistants for Mathematics Exam Grading: Reliability, and Practical Usab…
MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models
Expander Sparse Autoencoders: Parameter-Efficient Dictionaries for Mechanistic Interpreta…
Structuring the Space of Sociotechnical Alignment
AI Assistance for Human Review of Default Judgments
EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments
Online Safety Monitoring for LLMs
Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity
Distributed Attacks in Persistent-State AI Control
Safeguarding LLM Agents from Misalignment through Provenance Analysis
Adoption and Impact of Command-Line AI Coding Agents: A Study of Microsoft's Early 2026 R…
A General Neural Backbone for Mixed-Integer Linear Optimization via Dual Attention
Ophiuchus: Incentivizing Tool-augmented "Think with Images" for Joint Medical Segmentatio…