Fast A/B/n Testing: Exact Multi-Policy Comparison via Tree-Coupled Feedback Sharing
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
AQuA: Recursively Self-Improving Quantitative Trading Research Agents
BrainWAM: Action-Space Coordination of Semantic Priors and Predictive Dynamics for Autono…
SPARED: Reasoning-Based AI-Generated Image Detection via Adversarially Edited Data
NaviDC-OCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
EGRL: Edge generation-guided relation-aware learning for RNA-protein interaction predicti…
SteerBench-Work: A Benchmark for Agent Steering at Action Boundaries
Lines and Ladders: A Context-Aware Multi-Agent Framework for Large-Scale Retail Price Tax…
Designing AI Pipelines for Decision-Ready ITSM Intelligence
Multi-Agent Scheduling with LLM-Assisted Contract Net Negotiation for Stream Processing i…
Learning to Adapt Cross-Domain Preferences via Meta-LoRA for LLM Personalization
Research Assistant: AstraZeneca's Agentic System for R&D
Dual-Flow Transformers: Decoupling the Primary Prefill Path from Additional Decode Comput…
Sign Language Video Synthesis via Loss-Guided Multi-Expert GANs
$\varepsilon$-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution
Large Language Models Can Follow Instructions, But Not Many at Once: Phase Transitions in…
MindMemOS: A Portable and Self-Evolving Memory Operating Layer for AI Agents
Governed Persistent Memory: Source-Bound State Semantics and Fail-Closed Release for Long…
DiG-bench: Discovery in Games
CAS: A Causal Attribution Score for Local and Global Explainable Artificial Intelligence
@skills: Attention is all you have
Dead text or binding clause? Measuring and restoring constraint influence in black-box LL…
On the Expressive Power of Transformers
The Role of Natural Language Understanding in Multimodal Video-Based Dengue Diagnosis
Spatial Memory Agent: Experience-Grounded Procedure Memory for Spatial Intelligence
PROVE-RT: Generating Mechanized Theorem Prover Scripts for Real-Time Systems using LLMs
Practice Makes Unsafe: Skill Misevolution in Self-Improving LLM Agents
ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs
Beyond Retrieval: Query-Conditioned Reuse of Long-Horizon Agent Trajectories
Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs