Explorar

Noticias de IA

30934 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
JUMP: Single-Pass Membership Inference on Fine-Tuned Diffusion Language Models
arXiv cs.AI Research & Papers
RAIL Guard: Closing the Evaluation-to-Remediation Gap in Responsible AI for LLM Agents
arXiv cs.AI Research & Papers
Interactive Task Alignment as a POMDP
arXiv cs.AI Research & Papers
When to Plan: Learning to Select Between Reactive Control and Deliberative Planning
arXiv cs.AI Research & Papers
Mechanistic Attention Guidance for Agent Memory Refinement
arXiv cs.AI Research & Papers
ProEvent: An Event-centric Benchmark for Proactive Agents
arXiv cs.AI Research & Papers
Boundary-Seeking GAN-Augmented TabTransformer for Adversarially Robust Intrusion Detection
arXiv cs.AI Research & Papers
The Autonomous Agency Scale: A Behavioral Framework for Measuring Self-Directed Behavior …
arXiv cs.AI Research & Papers
KernelBench-Verified: Do LLM-Generated Kernels Actually Beat PyTorch?
arXiv cs.AI Research & Papers
Learning Structural Manipulability in Gate-Level Netlists Using Graph Neural Networks
arXiv cs.AI Research & Papers
CADENCE: Closing the Reasoning Gap via Coverage-Adaptive On-Policy Distillation
arXiv cs.AI Research & Papers
Benchmarking Machine Learning Models for Multi-Omics-Based Breast Cancer Prediction
arXiv cs.AI Research & Papers
Evidence Interfaces Shape How Retrieval-Augmented Readers Use Support
arXiv cs.AI Research & Papers
Self-Modifying Lean Proof Agents with Verifier-Grounded Benchmark Coevolution
arXiv cs.AI Research & Papers
AEC-DS: Adaptive Erasure Coding with PDP-Triggered Reputation and QoS-Aware Migration for…
arXiv cs.AI Research & Papers
Why Does Feedback-Augmented Self-Distillation Fail to Improve Retrieval-Interleaved Searc…
arXiv cs.AI Research & Papers
ZifaMem: Structured Memory for Persona, Preference, and Emotional Continuity in AI Compan…
arXiv cs.AI Research & Papers
Do Maps Still Matter for Machines: Revisiting the Role of Choropleth Maps in Foundation M…
arXiv cs.AI Research & Papers
WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Foo…
arXiv cs.AI Research & Papers
DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth
arXiv cs.AI Research & Papers
Comparing Spectrogram Front-Ends for Abnormal Heart-Sound Detection with a Convolutional …
arXiv cs.AI Research & Papers
Neural Controlled Differential Equations for EMT-Level Surrogate Modeling of Grid-Forming…
arXiv cs.AI Research & Papers
Reducing Per-Sample Harm in Stochastic Optimization
arXiv cs.AI Research & Papers
PhysAgent: Reflective Agentic Physics Control for Physically Plausible Video Generation
arXiv cs.AI Research & Papers
Reliable Remediation Impact Prediction for Black-Box Security Ratings
arXiv cs.AI Research & Papers
A Survey on the Verification of Reinforcement Learning Policies
arXiv cs.AI Research & Papers
Accurate and Efficient Long-Term Memory for LLM Agents
arXiv cs.AI Research & Papers
PRISM: Multimodal Terrain Mapping for Rover Navigation in Unstructured Environments
arXiv cs.AI Research & Papers
AoA: Theorem Proving Agent over Abstract Syntax Tree of Redesigned Language
arXiv cs.AI Research & Papers
Symbolic Augmentation Closes a Canonical-Equivalence Blind Spot in Neural Fact-Checkers