Metanormative Theory for RL-Based Moral Agents
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
SuperLocalMemory 4.0: The Governed Memory Operating System for AI Agents
OBLIVION: Workflow-Level Operational Skill Unlearning for Deployed Agents
Exploring LLM Capabilities for Situational Understanding and COLREG compliance on real-wo…
Fair on the Surface? Benchmarking Hidden-Output Fairness Gaps in LLM Recommenders
Learning to Modulate, Not to Cycle: Soft Actor---Critic Recovers Inverter-Style Heat-Pump…
StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoni…
ToolVision: Learning When and How to Use Visual Tools with Capability-Aligned Supervision
CAP: A Scalable Benchmark for Evaluating Cross-Site Browser Agents with Complex Actions a…
Forgotten History or Test-of-Time? Retrospect and Prospect on RAG from an IR Perspective
RAG-Based Auto-Configuration for Industrial Fieldbus Devices
TRACE-Memory: Public-Conditioned Retrieval and Utility-Aware Evidence Admission for Perso…
Estimating Uncertainty in Galaxy Morphology Classification
Not Worth Another Token: Marginal Value Estimation for Efficient Deep Research Agents
Automating Deception: Scalable Multi-Turn LLM Jailbreaks
CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark a…
Adversarial Latent-State Training for Robust Policies in Partially Observable Domains
Mendel G\"odel Machine: Recursive Self-Improving Coding Agents via Comparative Evolution
Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Comp…
NL2SHACL-Bench: A Benchmark Suite for Natural Language to SHACL Translation
A Sobering Look at Tabular Data Generation via Probabilistic Circuits
CliniCARE-Bench: Clinical Calibrated Audit of Medical Reasoning in EHR
SurgLAT: Surgical Latent Attention Tracking for Depth-Aware Robotic Laparoscope Control
Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy
Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control
ZhuLong: Execution-Grounded LLM Agent for EDA Scripting with Offline API Self-Exploration
REIN: Bridging the Gap between Reasoning and Reliability via Reflection and Abstention Al…
The Cost of Language: Centroid Erasure Exposes and Exploits Modal Competition in Multimod…
Decided Upstream, Written Late: Locating and Pricing the Cross-Lingual Refusal Circuit of…
Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?