The Approximation Rank of Softmax Attention: Sharp Geometric Laws and Robust Interaction …
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
LongPIBench: A Long-Context Benchmark for Prompt Injection
Low-Altitude Fluid Antenna Network with Multi-Agent Reinforcement Learning
RealSWE: A Compositional Evaluation of Coding Agents under Realistic User Requests
WeAgent-MMSearch: Native Text-Vision Interaction for Multimodal Search Agents
Quantization-Triggered Backdoors in Language Models: Cross-Quantizer Transferability and …
LandingAgent: A Reference-Annotated Dataset and Agentic Generation Framework for Landing …
SOMTab: Set-Order Mamba for Efficient Tabular In-Context Learning
FedEHR-Agents: Federated Agentic Optimization for Automated EHR Modeling
Actionable CBFI: Integrating Structural Decomposition and Causal Counterfactual Recourse …
Deriving Scaling Laws for OpenEuroLLM Models: Learning Rate, Batch Size and Loss
FISGuard: Defending Against Membership Inference via Fixed Input Subspaces
Anatomy-Aware Promptable Segmentation with Online Interactive Training for AUTOPET V
Layered LLM Defenses as an Ensemble: Access Tiers, Inference Cost, and the Measured Failu…
Beyond Task-Only Matching: Personalized Skill Routing with Counterfactual Evaluation
Embedding Models for Stance-Aware Argument Retrieval
A milestone in expanding access to AI
PACE: Publisher-Adaptive Content Extraction via Agentic Automation
The Effect of Emotional Context on Large Language Models' Endorsement of Premature Decisi…
From Documents to Reasoning: A Validated Synthetic Data Pipeline and Semantic-Aware Fine-…
Select, Don't Train: The Benefits of Modular Entity Disambiguation with LLM-Based Selecti…
Sledgehammer or Scalpel? A Fine-grained Adaptive Framework for Implicit Hate Speech
A comprehensive and trustworthy benchmark of AI methods for change detection in Earth obs…
SpikeOPD: Stable On-Policy Distillation for Autoregressive Spiking Language Models
LoopArena: Benchmarking Models as Runtime Controllers for Loop Engineering
Compositional Failure in Audio-Visual LLMs: Late-Layer Prior Dominance Under Cross-modal …
Evaluating Loss Functions in Differentiable Out-of-Domain Sound-Matching with Partial Par…
Finding Where the Buck Stops: An Automated Failure Attribution-Based Reflection Framework…
RECAST: Recent & Context-Aware Sampling for Test-Time Adaptation in Streaming Biosignals
Regime-Aware Portfolio Management via Retrieval-Augmented LLM-Guided Expert Switching