RAW
Dedicated GPU and CPU servers for AI workloads, provisioned in 3 seconds over a 100% REST API with flat pricing and zero egress fees.
At a Glance
Engagement
Available On
Alternatives
Listed Sep 2026
About RAW
RAW is a bare-metal cloud platform built specifically for AI inference, training, and agent workloads. Founded in 2025 by Michael Jakob — who previously built Eulerpool, a financial data platform — RAW grew out of a real $7,000 monthly cloud bill and the frustration of AWS complexity. The platform offers dedicated NVIDIA GPU and AMD EPYC CPU servers across five global regions, provisioned via a single REST API in as little as 3 seconds.
What It Is
RAW is a server-creation and fleet-scaling API designed to replace AWS, DigitalOcean, and GPU clouds for AI teams. Instead of navigating a 200-service console, users POST to a single endpoint to get a running server with full root SSH, CUDA, NVMe storage, and a public IP. The platform targets LLM inference, fine-tuning, AI agent hosting, and vector database deployments. RAW positions itself as "the anti-AWS console" — one API, flat pricing, and zero egress fees.
How the API Works
The entire product surface is a REST API with 24 endpoints. The workflow is intentionally minimal:
- Sign up with a single POST to get a Bearer token — no charge
- Add a card via
/billing/setup(Stripe-backed, $0 until deploy) - Deploy with
POST /deploypassingtypeandregion— server is live in ~13–30 seconds for CPU, 1–48 hours for GPU (dedicated hardware reservation) - Scale by repeating the same call; destroy with
DELETE /servers/:name
A CLI (npm install -g rawhq) and a web dashboard wrap the same endpoints. The dashboard provides 14 management tabs per server: metrics, networking, firewalls, volumes, backups, snapshots, rescale, rebuild, power, rescue, terminal, and an AI prompt helper for Claude/ChatGPT.
Infrastructure and Regions
RAW runs on dedicated hardware — no hypervisors, no shared tenancy — in Tier-3 datacenters across five regions:
- 🇩🇪 Frankfurt, Germany (
eu) - 🇮🇪 Dublin, Ireland (
eu-fi) - 🇺🇸 Ashburn, Virginia (
us) - 🇺🇸 Hillsboro, Oregon (
us-west) - 🇸🇬 Singapore (
sg)
CPU servers use AMD EPYC and Intel Xeon processors with NVMe SSDs. GPU servers use NVIDIA A100, L40S, and RTX cards with full CUDA. GPU SKUs are currently EU-only. All servers include unlimited bandwidth with no egress metering.
AI Stack Support
RAW is designed to run the full AI stack on dedicated hardware:
- LLM inference: vLLM, Ollama, TGI with OpenAI-compatible endpoints; supports Llama, Mistral, DeepSeek
- Training and fine-tuning: Full CUDA access, up to 96 GB VRAM, NVMe for checkpoints; LoRA or full training
- AI agents: Persistent, always-on CPU servers with root access; scale workers via API from CI
- Vector databases: Qdrant, pgvector, Milvus on dedicated NVMe
- Pre-install apps: Docker, PostgreSQL, Redis, Node.js, Nginx, Caddy, Python, Go, Bun, MySQL, OpenClaw
Deployment Model and Company Background
RAW is bootstrapped, headquartered in Munich, Germany, and describes itself as "AI-native" — product development, infrastructure automation, and operations are run primarily by AI systems rather than a large human team. The founder, Michael Jakob, migrated Eulerpool's entire stack from managed cloud to bare metal before building RAW as a product. The platform holds SOC 2 Type II certification and offers GDPR-compliant EU GPU regions. Per-second billing is available, and the 99.9% uptime SLA applies across all regions.
Community Discussions
Be the first to start a conversation about RAW
Share your experience with RAW, ask questions, or help others learn from your insights.
Pricing
2 vCPU · 4 GB (raw-5)
Entry-level CPU server with 2 vCPU, 4 GB RAM, 40 GB NVMe. EU only.
- 2 vCPU
- 4 GB RAM
- 40 GB NVMe
- Unlimited bandwidth
- EU regions (Germany & Ireland)
- Full root SSH
- 100% API
4 vCPU · 8 GB (raw-4x)
Most popular CPU server with 4 vCPU, 8 GB RAM, 80 GB NVMe. EU only.
- 4 vCPU
- 8 GB RAM
- 80 GB NVMe
- Unlimited bandwidth
- EU regions (Germany & Ireland)
- Full root SSH
- 100% API
8 vCPU · 16 GB (raw-8x)
Mid-range CPU server with 8 vCPU, 16 GB RAM, 160 GB NVMe. EU only.
- 8 vCPU
- 16 GB RAM
- 160 GB NVMe
- Unlimited bandwidth
- EU regions
- Full root SSH
16 vCPU · 32 GB (raw-16x)
Large CPU server with 16 vCPU, 32 GB RAM, 320 GB NVMe. EU only.
- 16 vCPU
- 32 GB RAM
- 320 GB NVMe
- Unlimited bandwidth
- EU regions
- Full root SSH
32 vCPU · 128 GB (raw-32d)
Dedicated RAM server with 32 vCPU, 128 GB RAM, 600 GB NVMe. All 5 regions.
- 32 vCPU
- 128 GB RAM
- 600 GB NVMe
- Unlimited bandwidth
- All 5 regions
- Full root SSH
48 vCPU · 192 GB (raw-48d)
Largest CPU server with 48 vCPU, 192 GB RAM, 960 GB NVMe. All 5 regions.
- 48 vCPU
- 192 GB RAM
- 960 GB NVMe
- Unlimited bandwidth
- All 5 regions
- Full root SSH
GPU 20 GB VRAM (raw-gpu-44)
Inference GPU server: 20 GB VRAM, 14 cores, 64 GB RAM, 3.8 TB NVMe. EU only. One-time setup fee applies.
- 20 GB VRAM NVIDIA GPU
- 14 CPU cores
- 64 GB RAM
- 3.8 TB NVMe
- Full CUDA
- EU only
- Supports Llama 8B, Whisper, Stable Diffusion
GPU 96 GB VRAM (raw-gpu-131)
Training GPU server: 96 GB VRAM, 24 cores, 256 GB RAM, 1.9 TB NVMe. EU only. One-time setup fee applies.
- 96 GB VRAM NVIDIA GPU
- 24 CPU cores
- 256 GB RAM
- 1.9 TB NVMe
- Full CUDA
- EU only
- Supports Llama 70B, custom training
GPU 96 GB VRAM Max (raw-gpu-131p)
Maximum GPU server: 96 GB VRAM, 24 cores, 768 GB RAM, 15 TB NVMe. EU only. One-time setup fee applies.
- 96 GB VRAM NVIDIA GPU
- 24 CPU cores
- 768 GB RAM
- 15 TB NVMe
- Full CUDA
- EU only
- Largest model configurations
Extra NVMe Volume
Persistent block storage (10 GB–10 TB) that survives rebuilds. Attach to any server.
- 10 GB to 10 TB
- Persists independently of server
- Survives rebuilds
- $0.07/GB/month
Extra IPv4
Additional public IPv4 address on any server.
- Additional public IPv4
- Assign and release from dashboard
Capabilities
Key Features
- 100% REST API with 24 endpoints
- GPU servers with full CUDA (NVIDIA A100, L40S, RTX)
- CPU servers from 2 vCPU to 48 vCPU
- 3-second server provisioning for CPU
- Full root SSH access on every server
- NVMe SSD storage on all servers
- Unlimited bandwidth with $0 egress
- 5 global regions (EU, US East, US West, Singapore)
- Per-second billing
- CLI (npx rawhq / npm install -g rawhq)
- Web dashboard with 14 management tabs
- Live metrics: CPU, RAM, disk, network
- Point-in-time snapshots and automated daily backups
- Persistent block storage volumes (10 GB–10 TB)
- SSH key management (import from GitHub)
- Web terminal (VNC console in browser)
- Pre-install apps: Docker, PostgreSQL, Redis, vLLM, Ollama, etc.
- Team access with Owner/Admin/Member roles
- SOC 2 Type II certified
- GDPR EU GPU regions
- Flat pricing — one price per server, no hidden fees
