Studying, Identifying, and Fixing Hidden Technical Debt in AI-Intensive Cyber-Physical Sy…
Explorar
Noticias de IA
29627 elementos — filtrados, clasificados y sin duplicados
Multimodal Auto-regressive Transformer Surrogate for Modeling Variable Operations and Qua…
V-FIND: Revealing the Intrinsic Forgery Knowledge Encoded in Video Forgery Detectors
SphUnc: Hyperspherical Uncertainty Decomposition and Causal Identification via Informatio…
TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluati…
FOUND-AF: Benchmarking ECG Foundation Models for Atrial Fibrillation Detection
SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation
Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent…
Standalone DINOv3 for Training-Free Open-Vocabulary Semantic Segmentation in Remote Sensi…
GraphCliff: Short-Long Range Gating for Modeling Critical Activity Changes Caused by Subt…
Secure AI Watermarking Framework for IP Protection in Multi-Tenant Cloud Platforms
CURV: Enhancing Chart Understanding Through Curriculum Visual Grounded Reasoning
Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC…
ADMITBench: A Safety-Governed Reference Framework for Evaluating the Admissibility of Ind…
Cortex: Compact Behavior Cloning for Quake with Frozen Visual Features
Assessing speech quality metrics for evaluation of neural audio codecs under clean speech…
MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Sc…
Output-Aware Rotation for INT2 KV-Cache Quantization
Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies an…
Your Agentic LLMs Secretly Encode Latent Signals of Indirect Prompt-Injection Exposure
Vulnerabilities, Secrets and Misconfiguration in the Highest-Exposure Docker Hub Images
Reachability Is Not Realization: Tracing the Sources of LLM Benchmark Gains
Measuring Explainer Stability via Attribution Separability
Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operato…
Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks
A neural operator framework for data-driven discovery of stability and receptivity in phy…
TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions
Getting the Parameters Right: A Difficulty-Graded Benchmark and Probe-Guided Training for…
Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling