Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Dr. Claw: An AI Scientist Workspace for Vibe Research
A Mathematical Framework for Legacy, Governance, and Decision Integrity in Enterprise AI
Reconstruct! Don't Encode: Self-Supervised Representation Reconstruction Loss for High-In…
Learning to Remember: End-to-End Training of Memory Agents for Long-Context Reasoning
EM^2Mem: Event-Centric Multimodal Memory for Large Language Models
Data-Driven Persona-Conditioned Agents for A/B Test Simulation
A Machine Learning-Driven Solution for Denoising Inertial Confinement Fusion Images
KItCAT: Knowledge Injection via Input Corruption for Auto-regressive Training
Runtime-Independent Persistent Agents: Preserving Identity, Memory, and Code Across Model…
Beyond Static Summarization: Proactive Memory Extraction for LLM Agents
Taming Modality Entanglement in Continual Audio-Visual Segmentation
InteractBench: Benchmarking LLMs on Competitive Programming under Unrevealed Information
TopoAlign: A Framework for Aligning Code to Math via Topological Decomposition
Are We There Yet? Assessing Computer-Use Agents for Blind Users' Accessible Interaction w…
Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis
APEX-EM: Non-Parametric Online Learning for Autonomous Agents via Structured Procedural-E…
AgentFactory: Towards Automated Agentic System Design and Optimization
Geometry-aware Latent Autoregressive Generative Model for PDEs in Complex Domains
Counterfactual Fragility Certificates: Exposing High-Confidence Brittleness under Structu…
SCALE:Scalable Conditional Atlas-Level Endpoint transport for virtual cell perturbation p…
Why Fine-Tuning Encourages Hallucinations and How to Fix It
Independent Reinforcement Learning in Discounted Markov Games
A Hybrid Insider Threat Detection Framework Combining Multi-Agent Simulation, Layered SIE…
Global Attention with Linear Complexity for Exascale Generative Data Assimilation in Eart…
What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal
When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluat…
OCGQuant: Outlier-Companion Grouping for NVFP4 Quantization
HBQ: Hierarchical Scaling Block Quantization with Hardware-Efficiency-Aware Design for Ac…
QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization