Where Does Toxicity Live? Mechanistic Localization and Targeted Suppression in Language M…
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
Prompt Optimization Is a Coin Flip: Diagnosing When It Helps in Compound AI Systems
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin
MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation
A Unified Framework for the Evaluation of LLM Agentic Capabilities
Periodic RoPE for Infinite Context LLMs
Do We Really Need Quantum Machine Learning?: A Multidimensional Empirical Study
SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Cont…
FPMoE: A Sparse Mixture-of-Experts Approach to Functional Code Generation
LoSATok: Low-dimensional Semantic-Acoustic Tokenizer for Cross-Domain Audio Understanding…
Revisiting Anthropomorphic Reflection Markers in Large Language Model Reasoning
Tensor Memory: Fixed-Size Recurrent State for Long-Horizon Transformers
Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdo…
Hallucination Behavior in Multimodal LLMs Across Agricultural Image Interpretation and Ge…
Detection Without Correction: A Two-Parameter Decomposition of Multi-Stage LLM Pipelines
Architecture-driven Shift: towards a lightweight selector for capturing the trends of log…
Do Models Know Why They Changed Their Mind? Interpretability and Faithfulness of Chain-of…
Case-Aware Medical Image Classification with Multimodal Knowledge Graphs and Reliability-…
LIFT and PLACE: A Simple, Stable, and Effective Knowledge Distillation Framework for Ligh…
Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillat…
GQLA: Group-Query Latent Attention for Hardware-Adaptive Large Language Model Decoding
Structured Belief State and the First Precision-Aware Benchmark for LLM Memory Retrieval
Quantifying the Reconstructability of Astrophysical Methods with Large Language Models an…
Compliance versus Sensibility: On the Reasoning Controllability in Large Language Models
S2MAM: Semi-supervised Meta Additive Model for Robust Estimation and Variable Selection
Negative Advantages Is a Double-Edged Sword: Calibrating advantages in GRPO for Search Ag…
When PCOS Meets Eating Disorders: An Explainable AI Approach to Detecting the Hidden Trip…
Early Decisions Matter: Proximity Bias and Initial Trajectory Shaping in Non-Autoregressi…
Speaking of Language: Reflections on Metalanguage Research in NLP