What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection un…
Explorar
Noticias de IA
21863 elementos — filtrados, clasificados y sin duplicados
Exact Symmetry as Algebra: A Machine-Verified Tensor Calculus that Enforces Physical Sele…
When benchmark inferences do not compose: Projectibility in AI evaluation
Facial-Expression-Aware Prompting for Empathetic LLM Tutoring
REAP: Automatic Curation of Coding Agent Benchmarks from Interactive Production Usage
From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for E…
Evidence-Ledger Adjudication for Claim-Evidence Traceability
On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment
Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating …
Can AI agents conduct open-ended AI research? Early evidence from two case studies
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adap…
Adaptively Robust LLM Monitoring via Activation Watermarking
Position: Evaluation Scores Are Perishable Knowledge Claims
CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games
Gated Adaptation for Continual Learning in Human Activity Recognition
Making Implicit Premises Explicit in Logical Understanding of Enthymemes
The Rise of AI in Weather and Climate Information and its Impact on Global Inequality
PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low…
DialectLLM: A Dialect-Aware Dialog[ue] Generation Framework Beyond Standard American Engl…
GBPP: Grasp-Aware Base Placement Prediction for Robots via Two-Stage Learning
When Should AI Follow? Task Structure and Joint Adaptation by Human and AI Agents
AI LEGO: Scaffolding Cross-Functional Collaboration in Industrial Responsible AI Practice…
Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homoge…
FPEdit: Robust LLM Fingerprinting through Localized Parameter Editing
Task and Skill Planning: Hierarchical Robot Planning with Black-Box Skills
Beyond Block Boundaries: Multi-Block Editing for Diffusion Large Language Models
The Cost of Knowing: A Resource-Aware Protocol for Benchmarking Hallucination Beyond Stat…
One-Frame Calibration with Siamese Network in Facial Action Unit Recognition
BioPro: Towards Difference-Aware Gender Fairness for Vision-Language Models
How does downsampling affect needle electromyography signals? A generalisable workflow fo…