Cost-Effective Repository Exploration for Agentic Issue Localization
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
MedSegBenchmarker: A Raw-Count-First Framework for Controlled 2D Medical Image Segmentati…
LLMODE: Aligning ODEs with LLMs via Gated Token Injection for Irregular Spatio-Temporal F…
AgenticRag-R1: Agentic Reinforcement Learning with Stack Memory for Multi-Step Reasoning,…
Forward-Deployed Full-Stack Engineering for Autonomous Cloud MLOps
Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformat…
Higher-Dimensional Rotary Position Embedding
SUP-MIMIC: A Multi-Task Clinical Diagnosis Benchmark for Evaluating LLMs' Robustness to C…
PhysWave: Physics-Guided Latent Diffusion Models for Controllable Spatial Audio Generation
Integrating adaptive human behavior into epidemic models with large language models
Argument-Aware Semantic Alignment of Normative Texts: A Toulmin-Based Neuro-Symbolic Appr…
SimGuide: Typed Multi-Context User Representations for Preference-Conditioned Agent Plann…
On the Prospects of Dynamic LLM Conversations in Software Development
Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems
Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level
Learning to Evaluate Before Improving: Automatic Rubric Induction for Automatic Research …
Denoising as Projection: Constrained Optimization with Gradient-Guided Diffusion
post-graph-rag: A PostgreSQL-Native Bi-Temporal Graph RAG Engine with Temporal Grounding …
Generalizable Multi-Agent Planning from Signal Temporal Logic Specifications via Diffusion
The Emergent Symbolic Structure of Artificial Neural Networks
Masked Distillation: Internalizing the Chain-of-Thought in Language Models
A Composition-Aware Pretraining Framework for Geospatial Foundation Models
On the Plasticity Collapse in Continual Machine Unlearning
Reference-Grafting Matches Fine-Tuning at Eliciting Sandbagged Capabilities
Benchmark Contamination: A Taxonomy Organized by Defeated Mitigation
Knowledge Distillation under Teacher Misspecification: An Order-Parameter Analysis of the…
BAITBENCH: Measuring Agent Reward Hacking with Optional Shortcuts Planted in ML Tasks
LCoT-GV: Graph Attention Networks for Verifying Long Reasoning Chains in Large Language M…
An Agentic Retrobiosynthesis Framework with Learned Frontier Selection
Beyond Dense States: Sparse Transcoders as Causally Testable Operators for LLM Latent Rea…