Self-Specialized Teachers for Domain Post-Training
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Spec2Twin-Chain: Orchestrating Bi-Level Optimization with LLMs for Blockchain Digital Twi…
Will the User Ever Know? Covert Indirect Prompt Injection Attacks on Tool-Using LLM Agents
Training-Free Action Correction for VLA Model Failures via Language Feedback
Balance of Benchmarks: Semantic Density Reweighting for Benchmark Multiplicity and Task-C…
TPR-Attention for Combinatorial Generalization
Beyond Fluency: A Rubric-Based Benchmark for Evaluating Saudi Dialect and Cultural Compet…
Spatial Matryoshka Training for Multi-Granularity Visual Document Retrieval
Rate-Coding Bundle Memory: A Unified Model of Memory and Control for Symbolic Computation…
IndicDetect: Evaluating Cross-Lingual LLM-Generated Text Detection for Hindi, Telugu, and…
Sleight of Word Benchmark: Can Language Models Notice If Their Own Output Was Tampered Wi…
LLMs Interpret, Embeddings Organize, Graphs Emerge: Agent-Driven Compilation of Scientifi…
Not Safe for All: Auditing the Dialect Penalty in Text-to-Image Safety Pipelines
Pak3H: Evaluating the Cost of Cultural Mismatch in LLM Alignment with a Human-Contextuali…
Interpretable Predictability-Based AI Text Detection: A Replication Study
When Less is More: Understanding When Token Filtering Helps and Fails in AI-generated Tex…
Reviving our data foundations is the most disruptive step to data maturity
Evaluating Tiny Recursive Models Across Training for Code Generation
Ideation Arena: Evaluating LLM Generated Research Ideas with Battle-style Human Expert As…
Toward Latent Language Model Skills Steering and Optimization: An Empirical Study
Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued …
Cross-Relational Preference Learning for Better LLM Instruction Following
Plant-Inspired AI: Plants as Inspiration for Novel Problem Formulations, and Two Case Stu…
TRACER: Per-Tool Context Retention for LLM Agents via Consequence-Attributed Reinforcemen…
Hallucination Mitigation for Large Vision-Language Models via Implicit Feature Stabilizat…
SynCrash: A Multi-Stage Pipeline for Zero-Shot Accident Detection and Localization in Tra…
Influence Is Not Authority: When Causal Guardrail Signals Make Legitimate Tool Use Look L…
REIGN: Refurbished Embeddings with Integrated Guidance Networks for Efficient Context-Len…
The reach of a verification tool decides its value: A controlled study of verification su…
Formal Concept Analysis with Three Types of Negation