An AI Scientist that Doesn't Drift: Taste, Structure, and Falsifiable Findings in a Quadr…
Explorar
Noticias de IA
21052 elementos — filtrados, clasificados y sin duplicados
The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes
The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI T…
ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents
Privacy-Preserving Data Drift Detection and Recovery for Large-Scale LLM Applications via…
TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Sourc…
Open-World Hierarchical Perception: Taxonomic Abstraction over Class-Agnostic Proposals f…
Do Evaluation Metrics Detect Errors in Classical Chinese to English Translations?
VTO: Visual Tool Orchestration for Video Anomaly Detection
When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains
Controlled Memory Interference in Continual LLM Agents
Two-Step MV-DeepONet: Probabilistic Operator Learning for Uncertainty Propagation Driven …
SiriusDeliver: Automating Data Warehouse Delivery at Tencent
Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification
Dynamic Coalition Formation and Communication Pricing in Skill-Based Agentic AI Systems
FreSH: Frequency-Segmented Hierarchical Multi-Expert Framework for Multivariate Time Seri…
Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinfo…
Frequency-Domain Dual-Branch Fusion for Medical Visual Question Answering
The Belief-Desire-Intention Ontology for modelling mental reality and agency
Defining Decentralization: An Ontological Perspective
Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization
MetaSpace: Metamorphic Testing for Spatial Cognition in Embodied Agents
From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Or…
Protecting patient privacy in clinical foundation models: Technical and legal perspectives
Adaptive Symmetry Discovery for Dynamical System Identification
Hierarchical Multi-Task Federated Learning in VANETs
Build it, Break it, Repeat: Benchmarking and improving LLM-manipulated disinformation det…
HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model …
$\texttt{DisMorph}$: learning to disentangle technical distortions from true biological c…
Emotion in an active inference model of human driving