SoK: AI-Augmented Binary Reversing
Explorar
Noticias de IA
30321 elementos — filtrados, clasificados y sin duplicados
Position: Coding Benchmarks Are Misaligned with Agentic Software Engineering
NarrativeWorldBench: A Frontier-Saturated Benchmark and a Latent World Model for Long-Hor…
Visuals Lie, Consistency Speaks: Disentangling Spatial Attention from Reliability in Visi…
The Price of Anarchy in Disaggregated Inference
An Evaluation of Data Leakage Risks in Tool-Using LLM Agents in Realistic Scenarios
Model Validation of Agentic AI Systems: A POMDP-Based Framework for Belief-State, Forecas…
MeiBRD: Meta-Learning Intraoperative Biomechanical Residual Deformation
Implicit vs. Explicit Prompting Strategies for LVLMs in Referential Communication
LegalHalluLens: Typed Hallucination Auditing and Calibrated Multi-Agent Debate for Trustw…
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipul…
C2FL: Clustered Continual Federated Learning under Spatial and Temporal Drift
Rethinking Cross-Layer Information Routing in Diffusion Transformers
Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety
LiveStarPro: Proactive Streaming Video Understanding with Hierarchical Memory for Long-Ho…
A Neuromorphic Trigger for Efficient Audio Event Detection
Models Take Notes at Prefill: KV Cache Can Be Editable and Composable
Surrogate Assisted Pedestrian Protection Design via a Foundation Model Orchestrated Workf…
Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learnin…
When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for O…
Do Large Language Models Always Tell The Same Stories?
DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Mod…
Structured Adversarial Camouflage via Voronoi Diagrams
Vision-language models for chest radiography do not always need the image
DRFLOW: A Deep Research Benchmark for Personalized Workflow Prediction
Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Ad…
From Paper to Program: Knowledge Externalization for AI-Assisted Quantum Many-Body Code G…
Membership Inference Attacks against Large Audio Language Models
DiffAttn: Diffusion-Based Drivers' Visual Attention Prediction with LLM-Enhanced Semantic…
Counterfactual Optimization of Baseball Pitch Sequences and Estimation of Its Impact on S…