gpu-cloud

(17)
Company
CoreWeave
Ai · Enterprise · Developer Tools

CoreWeave

The crypto-mining startup that became the picks-and-shovels supplier of the AI boom - renting Nvidia GPUs to the labs building the future, one megacluster at a time.

gpu-cloud · ai-infrastructureRead →
Company
DigitalOcean
Developer Tools · Saas · Enterprise

DigitalOcean

DigitalOcean built a public cloud by making servers feel less like a procurement exercise and more like a developer tool. Now it is trying to carry that same clarity into the expensive, unruly business of production AI.

digitalocean · cloud-computingRead →
Company
6Estates
Ai · Fintech · Enterprise

6Estates

6Estates is a Singapore-based enterprise AI company that turns messy, unstructured financial documents into structured decisions. Spun out of the NExT++ research center (NUS, Tsinghua, University of Southampton), it builds domain-specific large language models, intelligent document processing, and agentic AI tools that banks and lenders across Asia use to automate credit assessment, financial statement analysis, and trade finance checks.

6estates · document-aiRead →
Legend
Tycho Svoboda
Founder · Executive · Engineer

Tycho Svoboda

Tycho Svoboda is the co-founder and CEO of TensorPool, a Y Combinator Winter 2025 startup that lets machine learning engineers run training jobs on cloud GPUs with a single command-line tool, routing work to the cheapest available cloud and cutting compute costs by roughly half. He left Stanford with two credits shy of graduation to build it alongside two dorm-floor friends, joining a wave of young AI founders trading their degrees for the startup grind.

tycho-svoboda · tensorpoolRead →
Company
TensorPool
Ai · Developer Tools · Saas

TensorPool

TensorPool (YC W25) is a San Francisco startup that turns cloud GPUs into a git-style command line. Instead of wrestling with SSH sessions, provider dashboards, and idle-instance billing, machine learning engineers describe a training job in a config file and run 'tp job push' - TensorPool spins up a multi-node cluster across partner clouds, streams logs, saves outputs, and shuts everything down, billing only for runtime. It advertises GPU access at roughly half the cost of major cloud providers, with H100s starting around $1.99/hr. Founded by three Stanford dormmates, the company was accepted into Y Combinator's Winter 2025 batch.

gpu-cloud · on-demand-gpuRead →
Company
Modal
Ai · Developer Tools · Enterprise

Modal

Modal is a serverless cloud platform that lets developers run compute-intensive AI and data workloads - inference, training, fine-tuning, batch jobs, and sandboxed code - without managing infrastructure. Built from scratch in Rust with a custom container runtime, file system, and GPU memory snapshotting, it delivers sub-second cold starts and elastic autoscaling from zero to thousands of GPUs, billed by the second. Founded in 2021 by Erik Bernhardsson and Akshat Bubna, the New York-based company reached roughly $300M in annualized revenue and raised a $355M Series C in 2026 at a $4.65B valuation.

modal · modal-labsRead →
Company
Parasail
Ai · Developer Tools · Enterprise

Parasail

Parasail is a San Mateo-based AI inference cloud that lets developers run open-source and custom models without buying or committing to their own GPUs. Instead of owning chips, it orchestrates rented GPU capacity across roughly 40 data centers in 15 countries plus its own clusters, automatically routing each workload to the cheapest, fastest hardware. Founded in 2023 by Mike Henry (ex-Groq CPO, Mythic founder) and Tim Harris (Swift Navigation founder), the company processes on the order of 500 billion tokens a day and pitches a token-based economy - pay for output, not hardware.

ai-inference · gpu-cloudRead →
Company
VESSL AI
Ai · Developer Tools · Saas

VESSL AI

VESSL AI is an MLOps and GPU-cloud company that helps AI teams train, deploy, and scale machine learning models without wrestling with infrastructure. Its platform pools GPU capacity across on-premise clusters and multiple cloud providers, automatically routing workloads to the cheapest available resources - a hybrid, multi-cloud approach the company says can cut GPU spend by up to 80%. Founded in 2020 by ex-Google and PUBG engineers and split between Seoul and the San Francisco Bay Area, VESSL serves roughly 50 enterprise customers and 2,000-plus users, and has expanded into AI-agent tooling with its open-source Hyperpocket project.

mlops · llmopsRead →
Legend
Nikola Borisov
Founder · Engineer · Executive

Nikola Borisov

Nikola Borisov is the co-founder and CEO of Deep Infra, a Palo Alto cloud company that runs open-source AI models as a low-cost, low-latency inference service. A former competitive programmer who scaled messaging backends to hundreds of millions of users at imo.im and helped build HalloApp's founding backend, he bet early that inference, not training, would become the real bottleneck of the AI era. Deep Infra now processes close to five trillion tokens a week and raised a $107M Series B in May 2026 backed by NVIDIA, 500 Global, Felicis and others.

nikola-borisov · deep-infraRead →
Company
NodeShift
Ai · Developer Tools · Enterprise

NodeShift

NodeShift is a decentralized cloud platform that aggregates spare compute, GPU and storage capacity from independent data centers and web3 networks, then resells it through one simplified UI and a single API. By tapping underused infrastructure instead of building hyperscale data centers, NodeShift offers developers and enterprises GPUs, virtual machines and storage at prices it claims are up to 70-80% cheaper than traditional cloud providers - positioning itself as a 'sovereign AI cloud' for teams that want affordable, distributed, compliance-friendly compute.

cloud-computing · decentralized-cloudRead →
Company
MetalSoft
Enterprise · Saas · Developer Tools

MetalSoft

MetalSoft is a Chicago-based infrastructure software company that turns racks of physical servers, switches, and storage into cloud-like, self-service infrastructure. Its CloudPlex platform automates the full lifecycle of multi-vendor bare metal hardware - discovery, OS and firmware deployment, network fabric configuration, and storage - so enterprises and service providers can provision dedicated infrastructure on demand via a UI or API. Spun out of cloud hosting company Hostway in 2019 and led by serial founder Lucas Roh, MetalSoft raised a $16M Series A in December 2022 and has since pivoted hard toward orchestrating GPU clusters and 'AI factories' for telcos, data center operators, and Fortune 500 enterprises.

bare-metal · bare-metal-automationRead →
Company
Deep Infra Inc.
Ai · Developer Tools · Enterprise

Deep Infra Inc.

Deep Infra is a Palo Alto-based AI inference cloud that lets developers run hundreds of open-source machine learning models - large language models, image and video generation, speech, and embeddings - through a simple, OpenAI-compatible, pay-per-use API. The company owns and operates its own GPU fleet across multiple US data centers, processing trillions of tokens per week for companies that want production AI without managing infrastructure.

ai-inference · gpu-cloudRead →
Company
Hyperbolic
Ai · Developer Tools · Saas

Hyperbolic

Hyperbolic is an open-access AI cloud built for developers and researchers. It aggregates underutilized GPUs from data centers, mining farms and idle machines into an on-demand marketplace, and pairs that with serverless inference for state-of-the-art open-source models. The pitch is simple: fast, affordable access to compute and inference without sales calls or feature gates, at prices it says can run up to 75% below incumbents.

gpu-cloud · ai-infrastructureRead →
Company
GMI Cloud
Ai · Enterprise · Developer Tools

GMI Cloud

GMI Cloud is an AI-native GPU cloud built around NVIDIA's H100 and H200 accelerators. Founded in 2021 and headquartered in Mountain View, it sells on-demand and reserved GPU compute, an orchestration layer (Cluster Engine), and a low-latency Inference Engine to AI labs and enterprises building generative models. In 2025 NVIDIA named it one of seven Reference Platform Cloud Partners worldwide.

gpu-cloud · ai-infrastructureRead →
Legend
Alex Yeh
Founder · Executive · Operator

Alex Yeh

Alex Yeh is the Founder and CEO of GMI Cloud, a GPU-native AI cloud infrastructure company he built from Bitcoin mining data centers into a global AI infrastructure leader in just 30 days. GMI Cloud — one of only 6 NVIDIA Reference Platform Partners worldwide — raised $82M in Series A funding in 2024 and is behind a $12 billion sovereign AI infrastructure initiative in Japan. Yeh's mission: make building AI applications as simple as building a website on Shopify.

gpu-cloud · ai-infrastructureRead →
Legend
Haseeb Budhani
Founder · Executive · Operator

Haseeb Budhani

Haseeb Budhani is the Co-founder and CEO of Rafay Systems, a Sunnyvale-based infrastructure orchestration platform that simplifies Kubernetes and AI infrastructure lifecycle operations across public clouds, private data centers, and the edge. A serial entrepreneur with a track record of building and selling enterprise technology companies, Budhani previously co-founded Soha Systems - a secure remote access startup acquired by Akamai Technologies in 2016 - and has held senior product and engineering roles at Citrix, Infineta Systems, and others. Under his leadership, Rafay has raised $37.1M in total funding, built partnerships with Verizon, NVIDIA, and Cisco, and grown to serve major enterprise customers at the intersection of Kubernetes operations and AI infrastructure.

kubernetes · cloud-nativeRead →
Company
Together AI
Ai · Developer Tools · Enterprise

Together AI

Together AI is a San Francisco-based AI acceleration cloud that lets developers and enterprises train, fine-tune, and run open-source generative AI models on a high-performance GPU platform. Backed by $533M+ from General Catalyst, NVIDIA, Salesforce Ventures and others, it has become one of the most prominent challengers to the closed-model status quo.

generative-ai · open-source-aiRead →