Set-shifting Behavioral Test for Harnessed Agents
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search …
The Entanglement Wall: Activation-Space Probes as Risk Detectors, Not Context Adjudicators
Operational Evidence Gaps for LLMs in Fraud Detection and Trust-and-Safety Workflows
SingGuard-NSFA: Extensible Guardrails for Agentic AI via Generative Reasoning and Real-Ti…
Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Frame…
A Hybrid Mamba for Audio-Visual Navigation
Discourse-Aware Policy Analysis with Argumentation: A Hybrid LLM-Symbolic Framework for D…
Deconstructing Actor-Critic: A Large-scale Empirical Study of Design Components for Pract…
CAS I: A Geometric Coding Theorem
Tabular Foundation Models for Discrete Choice Estimation
SemaDiff: Identifying Semantic-Changing Commits with Generated Code and Tests
AI in Cyberpsychology: A systematic literature review of Cybersecurity enhancement by usi…
Active Beyond-Diagonal RIS Empowered Heterogeneous Edge Computing: A Distributional Reinf…
SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy
Evaluation Ability Does Not Imply Optimization Utility: LLM-as-a-Judge Signals in Closed-…
Classifying daily activities needs posture, reconstructing them needs motion
The Caf\'e in Amsterdam: When the Incumbent Becomes the Oracle
Is the Statistical Advantage Worth the Cost? An Empirical Comparison of KANs and MLPs for…
Data-Efficient Adaptation of LLMs via Attention Head Reweighting
Consensus as Privileged Context for Label-Free Self-Distillation
Discrete Diffusion Models: A Unified Framework from Tokenization to Generation
Explainable Artificial Intelligence for Anomaly Detection in Banking Transactions: An Int…
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
AgentCompass: A Unified Evaluation Infrastructure for Agent Capabilities
A Self-Evolving Agent for Longitudinal Personal Health Management
Generative Compilation: On-the-Fly Compiler Feedback as AI Generates Code
Multi-Expert Routing for Multi-Domain Low-Resource OCR: A Manchu Case Study
MASPRM: Multi-Agent System Process Reward Model