LaCache: Robust Semantic Caching for LLM Serving
Explorar
Noticias de IA
29655 elementos — filtrados, clasificados y sin duplicados
Beyond One Output: Visualizing and Comparing Distributions of Language Model Generations
SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents
GraRe: Grasp Candidate Re-Ranking for Frozen 6-DoF Grasp Detectors
Boosting Generalizable Depth Estimation in Endoscopy by Mixture of Lightweight Experts an…
DSETA: A Dual-Stage Continual Learning Framework for Travel Time Prediction in Dynamic Tr…
DAPD: Dual-Anchored Policy Distillation
DeVIT: Low-Power Vision Transformer Acceleration Using Delta Computation
When Extreme Darkness Meets Motion Blur: MeanFlow for Unified RAW Restoration
Context Compaction Theory
When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing
Mamba with Hierarchical Memory: Solving Representation Bottleneck in Long Sequence Modeli…
QARIMA: A Quantum Approach To Classical Time Series Analysis
Tensor Probabilistic Model Checking of Finite-Horizon Markov Chains (Extended Version)
Artificial Intelligence and Modeling & Simulation: An Overview
The Gate, Not the Cache: Gate Provenance Bounds the Closed-Loop Reliability of Training-F…
Tracing LLM Behavior to the Training Data with Empirical Next-Token Distributions
ORCA: ORgan-Centroid Aggregation for Training-Free 3D CT Visual Token Compression
Artificial Intelligence for the Characterization of Particles and Fibers by Optical Micro…
Sixteen models, fewer than two voices: measuring ensemble dispersion where no answer is u…
Verifiable Checks for Business Rule Consistency
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights …
AutoFOAM: The Self-Refining Autonomous OpenFOAM Agent
OrEdge: Efficient Multi-Modal Anomaly Detection in Distributed Software Systems via Ortho…
Exposed by Design: A Dynamic Security Assessment of Internet-Facing MCP Servers at Scale
Optimising for Flourishing: Flourishing Metrics and Return on Flourishing as Success Crit…
TALSC: Timeliness-Aware Large-Small VLM Collaboration for Infrastructure-Assisted Autonom…
Response Magnitude as a Dominant Signal for Held-Out CRISPRi Perturbation Effect Predicti…
FRAMES: Guarded and Dual-Objective Skill Evolution for Agents in Policy-Governed Enterpri…
Evidence-Unit Fairness and the Limits of Query-Adaptive Sparse-Dense Fusion in Financial …