# RunPod

> RunPod is an AI cloud infrastructure company that provides on-demand GPU compute for training, fine-tuning, and deploying AI/ML models. Founded in 2022 by two former Comcast engineers who pivoted their Ethereum mining rigs into AI servers, RunPod grew to $120M ARR with just $22M raised by early 2026, serving 500,000+ developers across 183 countries. Its marketplace model, per-second billing, and support for 30+ GPU SKUs — from consumer RTX 4090s to enterprise H100s and B200s — make it a capital-efficient disruptor to hyperscaler GPU clouds like AWS, GCP, and Azure.

- **Founded:** 2022
- **Headquarters:** Moorestown, New Jersey, USA
- **Founders:** Zhen Lu (Co-founder & CEO), Pardeep Singh (Co-founder & CTO)
- **Team size:** ~82–90 employees as of early 2026; remote-first with presence in San Francisco
- **Products:** GPU Pods, Serverless Endpoints, Instant Clusters, Serverless CPU, RunPod Hub
- **Notable:** $1M ARR within 9 months of launch (mid-2022), 100,000 developers on platform by May 2024, $20M seed round raised at minimal dilution with only ~$22M total raised

## Products & services

- **GPU Pods** — Persistent, dedicated GPU instances (like traditional VMs) with full control over OS, drivers, and container environment. Per-second billing. Supports 30+ GPU SKUs from RTX 4090 to H100 and B200.
- **Serverless Endpoints** — Auto-scaling inference endpoints with sub-500ms cold starts via FlashBoot technology. Ideal for production AI APIs with variable traffic.
- **Instant Clusters** — Multi-node distributed training clusters — up to 64 H100s provisioned within minutes. Launched March 2025.
- **Serverless CPU** — Extends platform beyond GPU to CPU workloads for data prep, agent orchestration, and backend processing.
- **RunPod Hub** — Marketplace for one-click deployment of open-source AI projects. Takes up to 7% commission on compute spend.
- **Network Volumes** — Persistent NVMe-backed storage starting at $0.05/GB/month.
- **AI Model APIs** — Pre-deployed AI models accessible via simple HTTP requests — no infrastructure setup needed. Launched August 2025.

## Achievements

- $1M ARR within 9 months of launch (mid-2022)
- 100,000 developers on platform by May 2024
- $20M seed round raised at minimal dilution with only ~$22M total raised
- $120M ARR as of January 2026 — 90% YoY growth
- 500,000+ developers across 183 countries
- 8 exabytes of annual network traffic processed
- 120% Net Dollar Retention — customers consistently expand usage
- Mid-60s to high-70s percent gross margins
- Published inaugural 2026 State of AI Report based on anonymized production platform data
- First GPU cloud to launch per-second billing at scale with no minimum commitment
- 155% YoY developer signup growth

## Latest updates

- **2026-03** — Published inaugural 2026 State of AI Report based on anonymized platform data from 183 countries — found Alibaba Qwen overtook Meta Llama as most deployed self-hosted LLM, and Nvidia B200 usage scaled 25x in 2025.
- **2026-01** — Announced $120M ARR milestone via TechCrunch exclusive; 500,000+ developers on platform; 90% YoY growth and 155% developer signup growth.
- **2025-12** — Launched Serverless CPU for non-GPU workloads (data prep, agent orchestration, backend processing).
- **2025-08** — Launched Public Endpoints — ready-to-use AI models accessible via HTTP API with no infrastructure setup.
- **2025-03** — Launched Instant Clusters — provision 16–64 H100s in minutes for distributed multi-node training.
- **2024-05** — Raised $20M seed round co-led by Intel Capital and Dell Technologies Capital; reached 100,000 developer milestone.

## Links

- Website: https://runpod.io
- LinkedIn: https://www.linkedin.com/company/runpod-io
- Twitter/X: https://twitter.com/runpod_io
- GitHub: https://github.com/runpod

---

Profile page: https://yespress.io/runpod
Published by YesPress — https://yespress.io
Last updated: 2026-04-02
