AgentAntibody: An Adaptive Immune System for Defending LLM Agents against Prompt Injection
Explorar
Noticias de IA
29349 elementos — filtrados, clasificados y sin duplicados
Hardware Design and Security in the Era of Chiplets and LLMs
Teaching Nemotron Greek: Mining a Corpus, Adapting Retrieval, and Grounding Generation fo…
Reconstructing Persistent Worlds from Narratives for Narrative-Grounded Interactive Exper…
Calibrating Transformer Attention via Task-Space Sensitivity Feedback
NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Le…
CheckOne: Lightweight Fault Detection and Mitigation for Vision Transformers
Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings
When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LL…
FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents
MarsCast: Transfer Learning of AI Weather Foundation Models to Planetary Atmospheres
SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving
Stabilizing Multi-Attack Adversarial Training via Bandit Optimization
Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs
FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Profe…
The LLM Proposes, the Executive Disposes: A Self-Verifying Agent Instrument that Dissocia…
MediRec: Enhancing Chinese Medication Recommendation with Explainable Clinical Reasoning
Beyond Linear Dynamics: Neural Bilinear Dynamical Models for Time Series Forecasting
Provable Limits and Certified Deferral for Verbalized Uncertainty in Small Language Models
Text2GraphQuery-Bench: A Text to Graph Query Benchmark
Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits
Revealed Rationality: Label-Free Evaluation and Regularization from Representation Theore…
Compass: Continuously Aligning Social Media Feeds via In-Situ Reflections
The First EgoCross Challenge at EgoVis 2026: Cross-Domain Egocentric Video Question Answe…
Topology-Aware Reasoning over Incomplete Knowledge Graph with Graph-Based Soft Prompting
Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation
Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Cl…
ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment
AgentForge: An Immersive Role-Playing Platform for Learning Agentic Software Engineering
Beyond the QBER Threshold: A Temporal QBER Based Machine Learning Framework for Multi Att…