Knowing What Not to Answer: Selective Non-Compliance in Vision-Language Models
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Building a research-software catalog with a coding agent: from hackathon prototype to pub…
Refuse without Refusal: A Structural Analysis of Safety-Tuning Responses for Reducing Fal…
Persistent Teacher Anchoring for Tool-Using Agents
Training-Free Halving of Activated Experts in Fine-Grained Mixture-of-Experts Models
Dynamic Adaptation of the LLM Context for Generating Routines with Coupled Semantics
When Do Internal Probes Beat Reading the Answer? Miscalibrated Readouts and Behavior-Conc…
Tracing Audio Grounding and Answer Selection in Audio LLMs
Adaptation Interfaces for In-Context Tabular Foundation Models in Time-to-Event Prediction
Atlas: Optimizing Deployment of Compound AI Workflows on Heterogeneous Clusters
Repeat-After-Me: Black-Box Adaptive Visual Prompt Injection
Beyond Code Generation: Reliability, Verification, and Cost Economics in the Agentic Soft…
Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based F…
Evidence Integration in Large Language Models
Abstraction Agent
What Moves? Localized Motion Representations for Compositional Scene Control
Necessary or Sufficient? Evaluating LLM Explanations With Behavioural Evidence
A Deep Generative Model for Synthesizing Labeled Wireless Signals
Trace2Tower: Transition-Aware EigenTrace Induction of Multi-Level Skills for LLM Agents
AlcaTRAz - Anchored Tree-Rule Defense Against Jailbreaks
Scalable Context Orchestration for Serving LLMs Over Voice
Don't Drop Dropout: Optimizing Layer Sparsity for Efficient LLM Training and Inference
Commonsense Reasoning in Computer Vision: Foundations, Recent Advancements, and Future Di…
A Systematic Evaluation of Cross-Lingual Consistency Enhancement Methods in Multilingual …
ACE: Adaptive Calibration-Free Expert Skipping for MoE-based LLMs
Moral Competence Before Moral Content: Why LLM Agents Lack the Prerequisites for Coherent…
TROVE: Adaptive Agent Skill Orchestration via Trace-Grounded Route Validation and Editing
CABAL: Multi-Agent Simulacra for Tracing the Effects of Collusive Bidding in Peer Review
Language models judge war differently when tested for alignment
Substrate-Aware AI Agents: Execution Context as a First-Class Input