Explorar

Noticias de IA

30934 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Set-shifting Behavioral Test for Harnessed Agents
arXiv cs.AI Research & Papers
LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search …
arXiv cs.AI Research & Papers
The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators
arXiv cs.AI Research & Papers
Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows
arXiv cs.AI Research & Papers
SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Ti…
arXiv cs.AI Research & Papers
Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Frame…
arXiv cs.AI Research & Papers
A Hybrid Mamba for Audio-Visual Navigation
arXiv cs.AI Research & Papers
Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for D…
arXiv cs.AI Research & Papers
Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Pract…
arXiv cs.AI Research & Papers
CAS I: A Geometric Coding Theorem
arXiv cs.AI Research & Papers
Tabular Foundation Models for Discrete Choice Estimation
arXiv cs.AI Research & Papers
SemaDiff: Identifying Semantic-Changing Commits with Generated Code and Tests
arXiv cs.AI Research & Papers
AI in Cyberpsychology: A systematic literature review of Cybersecurity enhancement by usi…
arXiv cs.AI Research & Papers
Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinf…
arXiv cs.AI Research & Papers
SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy
arXiv cs.AI Research & Papers
Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-…
arXiv cs.AI Research & Papers
Classifying daily activities needs posture, reconstructing them needs motion
arXiv cs.AI Research & Papers
The Caf\'e in Amsterdam: When the Incumbent Becomes the Oracle
arXiv cs.AI Research & Papers
Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for…
arXiv cs.AI Research & Papers
Data-Efficient Adaptation of LLMs via Attention Head Reweighting
arXiv cs.AI Research & Papers
Consensus as Privileged Context for Label-Free Self-Distillation
arXiv cs.AI Research & Papers
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation
arXiv cs.AI Research & Papers
Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Int…
arXiv cs.AI Research & Papers
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
arXiv cs.AI Research & Papers
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
arXiv cs.AI Research & Papers
AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities
arXiv cs.AI Research & Papers
A Self-Evolving Agent for Longitudinal Personal Health Management
arXiv cs.AI Research & Papers
Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code
arXiv cs.AI Research & Papers
Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study
arXiv cs.AI Research & Papers
MASPRM: Multi-Agent System Process Reward Model