Towards Faithful Agentic XAI: A Verification Method and an Open-World Benchmark for Bette…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Singular Vectors of Attention Heads Align with Features
Dr-CiK: A Testbed for Foresight-Driven Agents
DiagramRAG: A Lightweight Framework to Retrieve Scientific Diagram for Figure Generation
The Shape of Overthinking: Backtracking Bursts in Long Reasoning Traces
Geometry of Human Perceptual Domains Emerges Transiently in LLM Representations
Do Agents Think Deeper? A Mechanistic Investigation of Layer-Wise Dynamics in Sequential …
Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows
Show, Don't TELL: Explainable AI-Generated Text Detection
SKILLC: Learning Autonomous Skill Internalization in LLM Agents via Contrastive Credit As…
LASER: Learning Active Sensing for Continuum Field Reconstruction
Compositional Consistency-Guided Decoding for Three-Way Logical Question Answering
PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Manageme…
AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models
Causal Direct Preference Optimization for Distributionally Robust Generative Recommendati…
When Context Flips, Safety Breaks: Diagnosing Brittle Safety in Aligned Language Models
Revealing Algorithmic Deductive Circuits for Logical Reasoning
Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches
Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems
The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling
The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic V…
Differential syntactic and semantic encoding in LLMs
A Policy-Driven Runtime Layer for Agentic LLM Serving
JMedEthicBench: A Multi-Turn Conversational Benchmark for Evaluating Medical Safety in Ja…
Cyberbullying Governance on Social Media: A Unified Framework from Content Identification…
RULER: Representation-Level Verification of Machine Unlearning
HGMEM: Hypergraph-based Working Memory to Improve Multi-step RAG for Long-Context Complex…
Optimal and Diffusion Transports in Machine Learning
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
Behavioural Analysis of Alignment Faking