F(AI)2R: Who Did What, and Who Checked? Verifiable AI Provenance as an Executable Skill
Explorar
Noticias de IA
30585 elementos — filtrados, clasificados y sin duplicados
KQFuzz: Knowledge-Guided Fuzzing for Quantum Libraries via Large Language Models
CaRE Compute-aware Remasking Evaluation Protocol for Masked Diffusion Language Models
Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels
Beyond Memory: A Templated Substrate for Heterogeneous Collaborative Knowledge Work with …
CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition
Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI…
MyMentorLLM: A psychotherapy GenAI environment with multimodal voice/text patients, train…
Seen, Said, or Forgotten? A Causal Audit of Visual KV Memory Across Dialog Turns
OmniPhys: Knowledge-Graph-Driven Benchmarking and Collective Optimization for Physical Co…
Emergent Latent-State Computation under Stochastic Volatility
Rashomon Alignment
Raven: High-Recall Sequence Modeling with Sparse Memory Routing
Tools Are Not Islands: Set-Level Tool Retrieval for LLM Agents via Query-Conditioned Hype…
Physics-Informed Neural Operator for Warm-Starting Background-Decomposed and Precondition…
DynaBridge: Dynamic Summary-Guided Cross-Task Multimodal Fusion for DASS-Structured Menta…
Image Quality Dependent Degradation for AI Systems
OrganLens: Organ-Specific Representation Learning for CT Foundation Models
SpectONet: A Physics-Guided Spectral Deep Operator Network for Euler-Bernoulli Beam Dynam…
Lowering the implementation barrier of neutral-atom quantum computing with agentic workfl…
Learning from 53.6K Real-World Developer Edits of AI-Generated Code
DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Mul…
Construction-Driven Injection: Linguistically-Grounded Edit-Based Code-Mixing Fingerprint…
Physics-Informed Broad Learning System: An Efficient Backpropagation-Free Framework for S…
How Do LLMs Read Bug Reports? An Empirical Study of Attention in LLMs for Automated Progr…
Interactive Reward Agent: GUI Task Evaluation via Environment-State Verification
Distributing Security Controls Through Harness Engineering
HiSkill: Empowering LLM Agents with Hierarchical Skill Graphs
Speculate While You Reason: Teaching Agents to Predict Their Next Tool Call via Joint Age…
Nudging Sustainable Choices through LLM-Generated Recommendation Explanations