Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
arXiv cs.AI Research & Papers
Intelligent Detection and Mitigation of Carpet-Bombing DDoS Attacks in SDN Using Retrieva…
arXiv cs.AI Research & Papers
APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented…
arXiv cs.AI Research & Papers
Decoupled Delay Compensation: Enhancing Pre-trained MARL Policies via Learned Dynamics Fi…
arXiv cs.AI Research & Papers
VesselSim: learning 3D blood vessel segmentation without expert annotations
arXiv cs.AI Research & Papers
Unified Neural Scaling Laws
arXiv cs.AI Research & Papers
AgentSociety: Incentivizing Agentic Social Intelligence
arXiv cs.AI Research & Papers
FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives
arXiv cs.AI Research & Papers
Tool Calling is Linearly Readable and Steerable in Language Models
arXiv cs.AI Research & Papers
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
arXiv cs.AI Research & Papers
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
arXiv cs.AI Research & Papers
Detached Skip-Links and $R$-Probe: Decoupling Feature Aggregation from Gradient Propagati…
arXiv cs.AI Research & Papers
When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of…
arXiv cs.AI Research & Papers
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
arXiv cs.AI Research & Papers
Co-folding model guided by structural proteomics
arXiv cs.AI Research & Papers
HRVConformer: Neonatal Hypoxic-Ischemic Encephalopathy Classification from the Heart Rate…
arXiv cs.AI Research & Papers
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
arXiv cs.AI Research & Papers
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
arXiv cs.AI Research & Papers
SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository …
arXiv cs.AI Research & Papers
The ATOM Report: Measuring the Open Language Model Ecosystem
arXiv cs.AI Research & Papers
From Static Context to Calibrated Interactive RL: Mitigating Distribution Shift in Multi-…
arXiv cs.AI Research & Papers
PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Pr…
arXiv cs.AI Research & Papers
MobileExplorer: Accelerating On-Device Inference for Mobile GUI Agents via Online Explora…
arXiv cs.AI Research & Papers
MedGuideX: Internalizing Decision Logic from Executable Guidelines into Large Language Mo…
arXiv cs.AI Research & Papers
UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems
arXiv cs.AI Research & Papers
Tail-Aware HiFloat4: W4A4 Post-Training Quantization for Wan2.2
arXiv cs.AI Research & Papers
The Attribution Blind Spot: Detecting When Language Models Rely on Memory Rather Than Ret…
arXiv cs.AI Research & Papers
It's Not the Capability: Harness Sensitivity Is Non-Monotone Across LLM Agent Tiers
arXiv cs.AI Research & Papers
Towards Feedback-to-Plan Decisions for Self-Evolving LLM Agents in CUDA Kernel Generation
arXiv cs.AI Research & Papers
Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Mu…