Logos: An Agent Harness on a Cross-Process Bus
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
A Deep Learning-Based Stacking Ensemble Framework for Turbofan Engine Remaining Useful Li…
DAMP: Decay-Aware Mixed-Precision Recurrent-State Quantization
Trajectory-Level Speculative Decoding for Diffusion Language Models
Curvature-Aware Radius Shrinkage for Adaptive Nearest Neighbor Classification
KLOD: Locality-Preserving Knowledge Editing via Non-Target Distribution Preservation
An Empirical Evaluation of Cross-City POI Recommendation on a Large-Scale Benchmark
Depth-Aware Pothole Detection Using YOLO and RT-DETR at the Edge
From Uncertainty to Clinical Risk: Severity-Aware Conformal Planning for Interactive Medi…
openJiuwen: Beyond Static Harnesses for Long-Horizon Coding Agents
When Teacher Guidance Misleads: Reward-Aligned On-Policy Distillation
Thinking Costs Tokens: When More Structure is Worth the Price
When Evidence Shapes Collaboration: Knowledge-Conditioned Topology Generation for Multi-A…
Benchmarking General Mobile Assistants in Challenging Real-World Scenarios
HyQuant: Hybrid-Precision Quantization for LLM Attention
Context Localization for Generalized Level-Based Evaluation in Knowledge-Based Systems
Automated Analysis Framework for Multilingual Climate-Health Literature Based on Multi-Ag…
Hypothesize, Evaluate, Refine: A Scientific Agent for PDE Discovery with Unknown Spatial …
Coverage, Not Credit: Failure-Credit Routing of Zeroth-Order Perturbation Budgets Does No…
Quanta Perception as Probabilistic Events
A milestone in expanding access to AI
The Instability of Safety: How Random Seeds and Temperature Expose Inconsistent LLM Refus…
Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in …
Retrieving Relations, Detecting Fallacies: A RAG Approach to Political Debate Analysis
Quantization-Triggered Backdoors in Language Models: Cross-Quantizer Transferability and …
String: An Agentic OS Where Every App Is a Markdown File
Scientific Graphics Program Synthesis via Dual Self-Consistency Reinforcement Learning
Rating the Raters: Rasch Measurement Theory for LLM Evaluation
Time Capsule of Testable Human Knowledge: 41 Years of Jeopardy! in a Single Free Local Mo…
TraceML: An Empirical Analysis of Human-Agent Planning in Machine Learning Development