On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
On a joint simultaneous learning of relevant feature subsets and subspaces in regression-…
A Unified Benchmark of Deep Learning Models for Multi-task 3D Brain Tumor Segmentation fr…
SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Ac…
Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO
HenTwin: A Multimodal Digital Twin Framework for Longitudinal Biological State Monitoring…
COntExt: Towards Context-Aware Ontology Extension from Operational Metrics
metasignal: A Python Package for Comprehensive Metacognitive Analysis and Decision-Making
Hypergradient-based Bilevel Reinforcement Learning with Improved Sample Complexity
Evaluating Federated Pre-Training: On the Reliability of Downstream Fine-Tuning and Intri…
Benchmarks Are Not Monolithic: Sample-Level Auditing and Orchestration for LLM Evaluation
Stratified Negation in RDF Rules: A Correct Approach (Extended Version)
LAWFUL: Law-Aligned Witness for Faithful Use of Latents
WaiT for the Signal: Simple Frequency-Aware Flow-Matching
SCMA: Structure-Conditioned and Metal-Aware Flow Matching for CT Metal Artifact Reduction
SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Indus…
Towards White-Box Deep Wireless Sensing
Beyond Borrowed Histories: Person-Aligned User Simulation for Interactive Role-Playing Ev…
WitCert: Sound Runtime Risk Observability and Gating for KV-Cache Quantization
SE(3)-MeanFlow: Few-Step Protein Backbone Generation on Lie Groups
AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction
Self-Play Meets Skill Evolution: Self-Evolving Search Agents that Pose, Solve, and Rememb…
Beyond Retrieval: Analytic Memory for Multimodal Agents
Beyond Component Testing: Validating Agentic AI Systems
Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates
Don't Mix Rewards, Mix Policies: Policy Decomposition and Optimization for Multi-Reward RL
MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft
Benchmarking LLM Competence on Logical Inference over Probability Operators
LEX-EC: A Lexical Evidence-Channel Audit Framework for Zero-Shot LLM Personality Classifi…