Test-Time Training Undermines Safety Guardrails
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
How Far Will They Go? Red-Teaming Online Influence with Large Language Models
MedExpMem: Adapting Experience Memory for Differential Diagnosis
PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Nego…
Evaluating Large Language Models in a Complex Hidden Role Game
An AI-Driven Framework for Energy-Efficient Environmental Monitoring in Smart Cities Usin…
CP or DP? Why Not Both: A Case Study in the Partial Shop Scheduling Problem
EDGE-OPD: Internalizing Privileged Context with Evidence Guided On-Policy Distillation
Foundation Protocol: A Coordination Layer for Agentic Society
EVE-Agent: Evidence-Verifiable Self-Evolving Agents
Mediative Fuzzy Logic: From Type-1 Foundations to Type-2, Type-3 and Quantum Extensions
TingIS: Real-time Risk Event Discovery from Noisy Customer Incidents at Enterprise Scale
Tabular PDF Information Extraction with Local LLMs and Layout-Aware Parsing: A Reliabilit…
VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Red…
RMA: an Agentic System for Research-Level Mathematical Problems
NeuroNL2LTL: A Neurosymbolic Framework for Natural Language Translation of Linear Tempora…
BOHM: Zero-Cost Hierarchical Attribution for Compound AI Systems
Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling
ArcMark: Distortion-Free Multi-Byte LLM Watermark via Optimal Transport
Towards Generalization of Block Attention via Automatic Segmentation and Block Distillati…
ZipMoE: Efficient On-Device MoE Serving via Lossless Compression and Cache-Affinity Sched…
Patterns vs. Patients: Evaluating LLMs against Mental Health Professionals on Personality…
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Mani…
Controlled Personalization in Legacy Media Online Services: A Case Study in News Recommen…
A drone-based framework for coral habitat mapping via weakly supervised segmentation
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
Preisach Attention: A Hysteretic Model of Sequential Memory
Spectral-inspired Operator Learning with Limited Data and Unknown Physics
Representational Alignment with Chemical Induced Fit for Molecular Relational Learning
Model Spec Midtraining: Improving How Alignment Training Generalizes