Pendra

Run open-source AI models on your own GPUs or UK-hosted hardware, with one OpenAI-compatible API and zero data retention.

Verified API available Free tier
Quick facts
What is it Run open-source AI models on your own GPUs or UK-hosted hardware, with one OpenAI-compatible API and zero data retention.
Pricing Freemium — from £99/mo
Free tier Yes
Platform Web Application
API Yes
Best for Clinical document processing (summarise records, extract from discharge notes), Privileged document review in legal firms
Domain registered 2026

Data updated Aug. 5, 2026

What does Pendra do?

Pendra is a managed inference platform built for teams in regulated industries that need to use AI without handing sensitive data to a third party. Instead of sending prompts to a public cloud, you install a small worker on your own GPUs (or use Pendra's UK-hosted ones), pull open-weight models like Qwen, Llama, or DeepSeek onto it, and call them through a single API. The whole setup is designed so your data never leaves your control — no retention, no foreign jurisdiction, no exposure.

Key features

What makes it stand out
01
Install a worker on your own GPUs with one command, or use Pendra's managed UK hardware
02
Pull and swap open-weight models (Qwen, Llama, DeepSeek, Gemma, gpt-oss) from the console or CLI
03
OpenAI- and Anthropic-compatible endpoints with Python and Node.js SDKs — your existing code just works
04
Zero data retention by architecture, with end-to-end encryption available to seal prompts to your worker
05
Hybrid deployment: mix self-hosted workers and Pendra-managed GPUs behind a single API

Who is Pendra for?

Who benefits most from this tool
Clinical document processing (summarise records, extract from discharge notes)
Privileged document review in legal firms
Citizen data automation for public sector (classification, redaction, response)

Pricing

Free tier available — start without a credit card

Free

Free
  • Personal use only
  • 1 self-hosted Pendra Worker
  • OpenAI-compatible API endpoint
  • Zero data retention
  • Sovereign jurisdiction
  • Python and Node.js SDKs
  • Community documentation

Pro

£99.0/month

Everything in Free, plus:

  • Commercial use rights
  • Up to 5 self-hosted Pendra Workers
  • Enhanced usage analytics
  • Worker monitoring & webhook alerts
  • Standard DPA included
  • Priority support (email)
  • Request logging (optional)
  • Private inference (end-to-end encryption)

Enterprise

Custom

Everything in Pro, plus:

  • Pendra-managed GPU infrastructure
  • Unlimited self-hosted Workers
  • Automatic data masking and redaction
  • Granular audit logging
  • Custom DPA and DPIA support
  • Dedicated account manager
  • SLA with uptime guarantee
  • SSO and role-based access control
  • Onboarding and integration support

Trust & presence

Domain Domain registered 2026

Gallery

Click any image to enlarge

Alternatives in AI inference

Lambda Verified AI inference

Cloud platform that rents NVIDIA H100/B200/B300 GPUs for training and running AI models at scale

Inferless Verified AI inference

Serverless GPU platform for deploying machine learning models in minutes, with auto-scaling and pay-per-use pricing.

Superlinked Verified AI inference

Self-hosted AI inference engine for search and document processing — deploy models on your own cloud infrastructure.

RunInfra Verified AI inference

Optimize open-source AI models for production — benchmark engines, tune latency, and deploy on any GPU.

RunPod Verified AI inference

Cloud GPU platform for AI developers — deploy, train, and scale AI models with on-demand infrastructure

Top 100k site
vLLM Verified AI inference

High-throughput LLM inference engine for fast, memory-efficient AI model serving.

Top 100k site
Cerebras Verified AI inference

Cerebras’ third-generation wafer-scale engine (WSE-3) is the fastest AI processor on Earth. It surpasses all other processors in AI-optimized cores, memory speed, and on-chip fabric bandwidth.

Top 100k site
GPU Mart Verified AI inference

Rent dedicated GPU servers and VPS for AI, rendering, and LLM hosting, starting at $85/month.

Share X LinkedIn Telegram
Pendra Visit