gpu-cloud

(28)
Company
Northern Data bought the chips. Then came the hard part.
Ai · Enterprise · Hardware

Northern Data bought the chips. Then came the hard part.

A former Bitcoin miner built a European AI cloud with a hardware bill to match. Its more revealing achievement was learning how to keep the machines earning.

gpu-cloud · ai-infrastructureRead →
Company
io.net and the art of renting somebody else’s GPU
Ai · Crypto · Enterprise

io.net and the art of renting somebody else’s GPU

The AI compute shortage has an awkward companion: idle hardware. io.net wants to connect the two, but its most revealing lessons concern trust, incentives and the price of a finished job.

gpu-cloud · ai-infrastructureRead →
Company
Aarna.ml and the Missing Cloud Around the GPU
Ai · Saas · Enterprise

Aarna.ml and the Missing Cloud Around the GPU

A GPU is a machine. A cloud is a promise that the machine will be there when someone asks, properly isolated, metered and ready to work. Aarna.ml built the software between those two ideas.

gpu-cloud · gpu-as-a-serviceRead →
Legend
Akshat Bubna and the Art of Making Hard Problems Smaller
Founder · Engineer · Executive

Akshat Bubna and the Art of Making Hard Problems Smaller

Before Akshat Bubna helped build an AI cloud, he won an international programming gold medal. The habit that connects those chapters is simple: make a hard problem smaller, then keep going.

akshat-bubna · modal-labsRead →
Company
The $200,000 Answer to a Billion-Dollar Question
Ai · Enterprise · Developer Tools

The $200,000 Answer to a Billion-Dollar Question

HPC-AI Tech sells a way to make expensive AI work less expensively. Its open-source video model supplies the most interesting receipt: an 11-billion-parameter experiment built for a reported $200,000.

gpu-cloud · colossal-aiRead →
Legend
Zhen Lu Put the Cloud in His Basement
Founder · Engineer · Executive

Zhen Lu Put the Cloud in His Basement

A quantum chemist quit the classroom, learned to code, and found his next experiment in two New Jersey basements. The result became Runpod, a developer cloud shaped by Reddit feedback, improvised infrastructure, and an aversion to making builders think about GPUs.

zhen-lu · runpodRead →
Legend
Rutvik Patel and the Long Road from Word Games to GPU Clouds
Engineer · Operator

Rutvik Patel and the Long Road from Word Games to GPU Clouds

Before the GPUs came a word game, a chat room and a peer-to-peer file-sharing app. Rutvik Patel's public trail is a study in how small systems lessons accumulate into the work of building cloud infrastructure for more than a million developers.

rutvik-patel · runpodRead →
Legend
Chase Lochmiller Built His Company Where the Power Was
Founder · Executive · Operator

Chase Lochmiller Built His Company Where the Power Was

Before AI needed gigawatts, Chase Lochmiller learned to look for energy nobody else could use. The former quant and Everest climber turned that instinct into Crusoe - a company built around the stubborn physical limits of computing.

chase-lochmiller · crusoeRead →
Company
Aethir rented the GPUs. Now it wants to build the capacity.
Ai · Crypto · Enterprise

Aethir rented the GPUs. Now it wants to build the capacity.

The Singapore company turned distributed computing into a business serving AI and gaming. Its next bet is more physical: bringing new GPU capacity online where customers are already asking for it.

gpu-cloud · depinRead →
Legend
Mike Henry Is Building the Air-Traffic Control for AI Compute
Founder · Engineer · Executive

Mike Henry Is Building the Air-Traffic Control for AI Compute

After a decade spent trying to make one chip do more, the Parasail co-founder is taking a wider view: connect the scattered machines, hide the plumbing, and let builders ask for tokens instead of servers.

mike-henry · parasailRead →
Legend
Jaeman An Is Building the Plumbing for AI's Expensive New Age
Founder · Engineer · Executive

Jaeman An Is Building the Plumbing for AI's Expensive New Age

Before AI could write the future, someone had to make the machines cooperate. The VESSL AI co-founder turned a decade of infrastructure headaches into a company built for the age of scarce GPUs and restless models.

jaeman-an · vessl-aiRead →
Company
CoreWeave
Ai · Enterprise · Developer Tools

CoreWeave

The crypto-mining startup that became the picks-and-shovels supplier of the AI boom - renting Nvidia GPUs to the labs building the future, one megacluster at a time.

gpu-cloud · ai-infrastructureRead →
Company
DigitalOcean
Developer Tools · Saas · Enterprise

DigitalOcean

DigitalOcean built a public cloud by making servers feel less like a procurement exercise and more like a developer tool. Now it is trying to carry that same clarity into the expensive, unruly business of production AI.

digitalocean · cloud-computingRead →
Company
6Estates
Ai · Fintech · Enterprise

6Estates

6Estates is a Singapore-based enterprise AI company that turns messy, unstructured financial documents into structured decisions. Spun out of the NExT++ research center (NUS, Tsinghua, University of Southampton), it builds domain-specific large language models, intelligent document processing, and agentic AI tools that banks and lenders across Asia use to automate credit assessment, financial statement analysis, and trade finance checks.

6estates · document-aiRead →
Legend
Tycho Svoboda
Founder · Executive · Engineer

Tycho Svoboda

Tycho Svoboda is the co-founder and CEO of TensorPool, a Y Combinator Winter 2025 startup that lets machine learning engineers run training jobs on cloud GPUs with a single command-line tool, routing work to the cheapest available cloud and cutting compute costs by roughly half. He left Stanford with two credits shy of graduation to build it alongside two dorm-floor friends, joining a wave of young AI founders trading their degrees for the startup grind.

tycho-svoboda · tensorpoolRead →
Company
TensorPool
Ai · Developer Tools · Saas

TensorPool

TensorPool (YC W25) is a San Francisco startup that turns cloud GPUs into a git-style command line. Instead of wrestling with SSH sessions, provider dashboards, and idle-instance billing, machine learning engineers describe a training job in a config file and run 'tp job push' - TensorPool spins up a multi-node cluster across partner clouds, streams logs, saves outputs, and shuts everything down, billing only for runtime. It advertises GPU access at roughly half the cost of major cloud providers, with H100s starting around $1.99/hr. Founded by three Stanford dormmates, the company was accepted into Y Combinator's Winter 2025 batch.

gpu-cloud · on-demand-gpuRead →
Company
Modal
Ai · Developer Tools · Enterprise

Modal

Modal is a serverless cloud platform that lets developers run compute-intensive AI and data workloads - inference, training, fine-tuning, batch jobs, and sandboxed code - without managing infrastructure. Built from scratch in Rust with a custom container runtime, file system, and GPU memory snapshotting, it delivers sub-second cold starts and elastic autoscaling from zero to thousands of GPUs, billed by the second. Founded in 2021 by Erik Bernhardsson and Akshat Bubna, the New York-based company reached roughly $300M in annualized revenue and raised a $355M Series C in 2026 at a $4.65B valuation.

modal · modal-labsRead →
Company
Parasail
Ai · Developer Tools · Enterprise

Parasail

Parasail is a San Mateo-based AI inference cloud that lets developers run open-source and custom models without buying or committing to their own GPUs. Instead of owning chips, it orchestrates rented GPU capacity across roughly 40 data centers in 15 countries plus its own clusters, automatically routing each workload to the cheapest, fastest hardware. Founded in 2023 by Mike Henry (ex-Groq CPO, Mythic founder) and Tim Harris (Swift Navigation founder), the company processes on the order of 500 billion tokens a day and pitches a token-based economy - pay for output, not hardware.

ai-inference · gpu-cloudRead →
Company
VESSL AI
Ai · Developer Tools · Saas

VESSL AI

VESSL AI is an MLOps and GPU-cloud company that helps AI teams train, deploy, and scale machine learning models without wrestling with infrastructure. Its platform pools GPU capacity across on-premise clusters and multiple cloud providers, automatically routing workloads to the cheapest available resources - a hybrid, multi-cloud approach the company says can cut GPU spend by up to 80%. Founded in 2020 by ex-Google and PUBG engineers and split between Seoul and the San Francisco Bay Area, VESSL serves roughly 50 enterprise customers and 2,000-plus users, and has expanded into AI-agent tooling with its open-source Hyperpocket project.

mlops · llmopsRead →
Legend
Nikola Borisov
Founder · Engineer · Executive

Nikola Borisov

Nikola Borisov is the co-founder and CEO of Deep Infra, a Palo Alto cloud company that runs open-source AI models as a low-cost, low-latency inference service. A former competitive programmer who scaled messaging backends to hundreds of millions of users at imo.im and helped build HalloApp's founding backend, he bet early that inference, not training, would become the real bottleneck of the AI era. Deep Infra now processes close to five trillion tokens a week and raised a $107M Series B in May 2026 backed by NVIDIA, 500 Global, Felicis and others.

nikola-borisov · deep-infraRead →
Company
NodeShift
Ai · Developer Tools · Enterprise

NodeShift

NodeShift is a decentralized cloud platform that aggregates spare compute, GPU and storage capacity from independent data centers and web3 networks, then resells it through one simplified UI and a single API. By tapping underused infrastructure instead of building hyperscale data centers, NodeShift offers developers and enterprises GPUs, virtual machines and storage at prices it claims are up to 70-80% cheaper than traditional cloud providers - positioning itself as a 'sovereign AI cloud' for teams that want affordable, distributed, compliance-friendly compute.

cloud-computing · decentralized-cloudRead →
Company
MetalSoft
Enterprise · Saas · Developer Tools

MetalSoft

MetalSoft is a Chicago-based infrastructure software company that turns racks of physical servers, switches, and storage into cloud-like, self-service infrastructure. Its CloudPlex platform automates the full lifecycle of multi-vendor bare metal hardware - discovery, OS and firmware deployment, network fabric configuration, and storage - so enterprises and service providers can provision dedicated infrastructure on demand via a UI or API. Spun out of cloud hosting company Hostway in 2019 and led by serial founder Lucas Roh, MetalSoft raised a $16M Series A in December 2022 and has since pivoted hard toward orchestrating GPU clusters and 'AI factories' for telcos, data center operators, and Fortune 500 enterprises.

bare-metal · bare-metal-automationRead →
Company
Deep Infra Inc.
Ai · Developer Tools · Enterprise

Deep Infra Inc.

Deep Infra is a Palo Alto-based AI inference cloud that lets developers run hundreds of open-source machine learning models - large language models, image and video generation, speech, and embeddings - through a simple, OpenAI-compatible, pay-per-use API. The company owns and operates its own GPU fleet across multiple US data centers, processing trillions of tokens per week for companies that want production AI without managing infrastructure.

ai-inference · gpu-cloudRead →
Company
Hyperbolic
Ai · Developer Tools · Saas

Hyperbolic

Hyperbolic is an open-access AI cloud built for developers and researchers. It aggregates underutilized GPUs from data centers, mining farms and idle machines into an on-demand marketplace, and pairs that with serverless inference for state-of-the-art open-source models. The pitch is simple: fast, affordable access to compute and inference without sales calls or feature gates, at prices it says can run up to 75% below incumbents.

gpu-cloud · ai-infrastructureRead →
Company
GMI Cloud
Ai · Enterprise · Developer Tools

GMI Cloud

GMI Cloud is an AI-native GPU cloud built around NVIDIA's H100 and H200 accelerators. Founded in 2021 and headquartered in Mountain View, it sells on-demand and reserved GPU compute, an orchestration layer (Cluster Engine), and a low-latency Inference Engine to AI labs and enterprises building generative models. In 2025 NVIDIA named it one of seven Reference Platform Cloud Partners worldwide.

gpu-cloud · ai-infrastructureRead →
Legend
Alex Yeh
Founder · Executive · Operator

Alex Yeh

Alex Yeh is the Founder and CEO of GMI Cloud, a GPU-native AI cloud infrastructure company he built from Bitcoin mining data centers into a global AI infrastructure leader in just 30 days. GMI Cloud — one of only 6 NVIDIA Reference Platform Partners worldwide — raised $82M in Series A funding in 2024 and is behind a $12 billion sovereign AI infrastructure initiative in Japan. Yeh's mission: make building AI applications as simple as building a website on Shopify.

gpu-cloud · ai-infrastructureRead →
Legend
Haseeb Budhani
Founder · Executive · Operator

Haseeb Budhani

Haseeb Budhani is the Co-founder and CEO of Rafay Systems, a Sunnyvale-based infrastructure orchestration platform that simplifies Kubernetes and AI infrastructure lifecycle operations across public clouds, private data centers, and the edge. A serial entrepreneur with a track record of building and selling enterprise technology companies, Budhani previously co-founded Soha Systems - a secure remote access startup acquired by Akamai Technologies in 2016 - and has held senior product and engineering roles at Citrix, Infineta Systems, and others. Under his leadership, Rafay has raised $37.1M in total funding, built partnerships with Verizon, NVIDIA, and Cisco, and grown to serve major enterprise customers at the intersection of Kubernetes operations and AI infrastructure.

kubernetes · cloud-nativeRead →
Company
Together AI
Ai · Developer Tools · Enterprise

Together AI

Together AI is a San Francisco-based AI acceleration cloud that lets developers and enterprises train, fine-tune, and run open-source generative AI models on a high-performance GPU platform. Backed by $533M+ from General Catalyst, NVIDIA, Salesforce Ventures and others, it has become one of the most prominent challengers to the closed-model status quo.

generative-ai · open-source-aiRead →