PhoenixRepair: Rethinking Repair Strategy Exploration in Software Agents
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Evaluating medical AI under missing information: same-provider judges and human raters ch…
AI Tour Meeting: Group Travel Planning by LLM Agents
AgentDebugX: An Open-Source Toolkit for Failure Observability, Attribution, and Recovery …
One Rewrite to Fix Them All? Type-Aware Repair Allocation for Text-to-Image Prompt Optimi…
Active Electrosensing and Communication in MARL-trained Weakly Electric Fish Collectives
Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States
Do AI-Native Biotechs Need Departments? Benchmarking Company World Models for AI-Driven D…
Semantic Primes as Explanans for Emotion in Large Language Models
AutoIndex: Learning Representation Programs for Retrieval
Functional Equivalence and Geometric Diversity in Neural Network Approximations: An Empir…
SciHazard: A Benchmark for Measuring Scientific Safety Risks with Decomposed Harm Scoring
ISO: An RLVR-Native Optimization Stack
Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editi…
When Does Machine Learning Beat Value Sorting? A Three-Dataset Diagnostic of Exposure-Wei…
Engineering Trustworthy Agentic AI for Critical Systems
Mitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learni…
LatentMT: Machine Translation with Latent Reasoning
MAGE: Human-Like Macro Placement via Agentic Multimodal Reasoning
From Operations to Elderly Care Outcomes: A Thematic Review of Industrial Engineering and…
Dual Adversarial Fine-tuning for Enhancing Robustness of Large Vision Language Model
SciCodePile: A 128GB Corpus and Executable Benchmark for Challenging Scientific Code Gene…
Saving the legacy of Hero Ibash: Evaluating Four Language Models for Aminoacian
ABOPD: Antibody CDR Design via On-Policy Distillation
SFGA: A Statistics-First Gating Architecture with Adjudicative Escalation for Trustworthy…
Neuro-Symbolic Meta-Policies for Temporal Knowledge-Graph Memory under Partial Observabil…
AlayaWorld: Interactive Long-Horizon World Modeling -- Full Technical Report
Operational Hallucination and Safety Drift in AI Agents
MIRA-Ev:A Benchmark for Granular Evidence Detection and Relational Reasoning in Clinical …
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU