Find Investable Startups and Competitors
Search thousands of startups using natural language—just describe what you're looking for
Top 50 Ai Gpu Cloud
Discover the top 50 Ai Gpu Cloud startups. Browse funding data, key metrics, and company insights. Average funding: $502.7M.
Sort by
McLean, United States
The startup provides a hybrid cloud platform that combines GPU cloud services with AI controller software for on-premises deployments, enabling enterprises to efficiently manage AI workloads. This solution allows businesses to optimize their data center resources while seamlessly scaling to the public cloud as needed.
Funding: $25.5M
Rough estimate of the amount of funding raised
Funding: $25.5M
Rough estimate of the amount of funding raised
Helsinki, Finland
Verda Cloud provides a full-stack GPU cloud platform optimized for AI workloads, offering on-demand instances, instant clusters, and serverless containers. The platform delivers high-performance compute, storage, and networking with a developer-first experience at significantly lower costs than hyperscalers. Users gain immediate access to cutting-edge NVIDIA hardware for efficient model training, experimentation, and scalable inference.
Funding: $18.7M
Rough estimate of the amount of funding raised
byFounders
byFounders
Funding: $18.7M
Rough estimate of the amount of funding raised
London, United Kingdom
Nscale provides a GPU cloud platform optimized for AI workloads, featuring on-demand compute and inference services, dedicated training clusters, and scalable GPU nodes. The platform addresses the high costs and inefficiencies associated with AI model training and deployment by offering a fully integrated infrastructure powered by renewable energy in Europe.
Funding: $3.3B
Rough estimate of the amount of funding raised
Sandton Capital Partners
Sandton Capital Partners
Funding: $3.3B
Rough estimate of the amount of funding raised
San Francisco, United States
Hyperbolic provides an open-access AI cloud that aggregates global GPU resources, enabling users to run AI inference and access compute power at significantly reduced costs. The platform addresses the high expenses associated with traditional cloud services by offering flexible, pay-as-you-go GPU access and opportunities for individuals and data centers to monetize idle machines.
Funding: $19M
Rough estimate of the amount of funding raised
PolychainVariant
PolychainVariant
Funding: $19M
Rough estimate of the amount of funding raised
Sharon AI provides scalable, GPU-accelerated cloud infrastructure for AI and High-Performance Computing (HPC) workloads. It offers on-demand access to a diverse fleet of high-performance GPUs, virtual servers, and cloud storage, enabling organizations to accelerate complex computations and AI development without significant capital expenditure.
Munich, Germany
Genesis Cloud provides a GPU cloud platform built on NVIDIA's reference architecture, delivering up to 35 times more performance for AI and machine learning workloads at 80% lower costs compared to traditional cloud providers. The platform ensures high security and compliance with EU regulations, enabling enterprises to efficiently manage and scale their AI applications.
Funding: $37M
Rough estimate of the amount of funding raised
Funding: $37M
Rough estimate of the amount of funding raised
London, United Kingdom
NexGen Cloud delivers fully managed AI superclouds and on‑demand GPU clusters powered by the latest NVIDIA hardware, including liquid‑cooled HGX H100/H200 and Blackwell GB200. The platform offers customizable compute configurations, managed Kubernetes or SLURM environments, and end‑to‑end MLOps support, with data residency in Europe and Canada and 100 % renewable‑energy operation.
Funding: $45M
Rough estimate of the amount of funding raised
Moore & Moore Investment
Moore & Moore Investment
Funding: $45M
Rough estimate of the amount of funding raised
Amsterdam, Netherlands
Provides a fully managed AI cloud platform powered by NVIDIA® H100 and H200 Tensor Core GPUs, offering scalable GPU clusters with InfiniBand networking for high-speed data processing. Enables efficient model training, fine-tuning, and inference with tools like MLflow, PostgreSQL, and Apache Spark, reducing the complexity and cost of deploying AI applications at scale.
Funding: $700M
Rough estimate of the amount of funding raised
Funding: $700M
Rough estimate of the amount of funding raised
Marietta, United States
The startup provides AI supercomputers that configure and host GPU servers optimized for deep learning and high-performance computing (HPC) applications. This offering enables organizations to scale their computational resources efficiently while minimizing the infrastructure costs associated with AI workloads.
Funding: $7.9M
Rough estimate of the amount of funding raised
Global Digital Holdings
Global Digital Holdings
Funding: $7.9M
Rough estimate of the amount of funding raised
Launceston, Australia
Firmus provides a modular AI infrastructure platform that combines liquid‑cooled, high‑density GPU clusters with a public AI cloud offering on‑demand and reserved GPU instances. Its AI FactoryOS orchestration layer automates workload placement, power and cooling management, and delivers real‑time telemetry, enabling AI labs and enterprise teams to train large models efficiently and predictably.
Funding: $10.5B
Rough estimate of the amount of funding raised
+ 3 Other investorsEllerston Capital
+ 3 Other investorsEllerston Capital
Funding: $10.5B
Rough estimate of the amount of funding raised
Singapore
Aethirs provides a decentralized cloud infrastructure that delivers on-demand access to enterprise-grade GPUs for AI model training and real-time gaming applications. This solution addresses the need for scalable, low-latency compute resources while ensuring high performance and security across a global network.
Funding: $24.8M
Rough estimate of the amount of funding raised
Funding: $24.8M
Rough estimate of the amount of funding raised
Mount Laurel, United States
RunPod is a cloud platform that provides globally distributed GPU resources for deploying and scaling machine learning applications, enabling developers to run AI workloads without managing infrastructure. The platform reduces cold-start times to under 250 milliseconds and offers flexible pricing, allowing users to efficiently handle fluctuating demand while minimizing operational costs.
Funding: $20M
Rough estimate of the amount of funding raised
Dell Technologies CapitalIntel Capital
Dell Technologies CapitalIntel Capital
Funding: $20M
Rough estimate of the amount of funding raised
Roseland, United States
CoreWeave provides an AI-native cloud platform built on next-generation infrastructure, purpose-built for complex AI workloads. The platform offers specialized GPU compute, storage, and high-performance networking within a Kubernetes-native environment. This specialized offering accelerates AI development cycles, training, and inference with enhanced efficiency and operational control.
Funding: $650M
Rough estimate of the amount of funding raised
Jane Street CapitalMagnetar Capital
Jane Street CapitalMagnetar Capital
Funding: $650M
Rough estimate of the amount of funding raised
San Francisco, United States
Together AI provides an AI-native cloud platform engineered for accelerating model training, fine-tuning, and inference on performance-optimized GPU infrastructure. The platform offers a comprehensive suite of tools, including a model library, serverless inference APIs, and self-service GPU clusters featuring frontier hardware. This infrastructure delivers industry-leading unit economics and performance for developers building large-scale generative AI applications.
Funding: $513.5M
Rough estimate of the amount of funding raised
Salesforce Ventures
Salesforce Ventures
Funding: $513.5M
Rough estimate of the amount of funding raised
San Jose, United States
GMI Cloud provides instant access to NVIDIA H100 GPUs for training and deploying generative AI applications, utilizing a Kubernetes-based cluster engine for efficient workload orchestration. This platform addresses the need for rapid GPU provisioning and management, enabling developers to focus on building AI models without the complexities of infrastructure setup.
Funding: $142M
Rough estimate of the amount of funding raised
Headline Asia (formerly Infinity Ventures)
Headline Asia (formerly Infinity Ventures)
Funding: $142M
Rough estimate of the amount of funding raised
West Palm Beach, United States
Vultr provides cloud infrastructure with dedicated clusters and on-demand virtual machines powered by AMD and NVIDIA GPUs, enabling efficient deployment of AI and high-performance computing workloads. The platform offers scalable solutions at competitive pricing, addressing the need for accessible and powerful computing resources for developers and businesses globally.
Funding: $333M
Rough estimate of the amount of funding raised
AMD VenturesLuminArx Capital Management LP
AMD VenturesLuminArx Capital Management LP
Funding: $333M
Rough estimate of the amount of funding raised
성남시, South Korea
Kakao Enterprise provides an integrated AI‑cloud platform that combines scalable compute, AI services, and search capabilities into a single environment. Its hybrid GPU‑as‑a‑Service offers on‑demand high‑performance GPU resources without upfront hardware costs, while a data‑centric console streamlines provisioning, monitoring, and cost optimization for enterprises and public organizations seeking to accelerate AI and digital transformation projects.
Funding: $175.8M
Rough estimate of the amount of funding raised
Funding: $175.8M
Rough estimate of the amount of funding raised
Cupertino, United States
Lepton (NVIDIA DGX Cloud Lepton) is a unified AI platform that aggregates a global network of GPU resources from NVIDIA Cloud Partners, cloud providers, and on‑premise environments into a single developer interface. It abstracts infrastructure details, letting AI developers prototype, train, and deploy models with consistent APIs, unified billing, and built‑in monitoring across multi‑cloud deployments.
Funding: $11M
Rough estimate of the amount of funding raised
Funding: $11M
Rough estimate of the amount of funding raised
San Francisco, United States
Lambda provides an on‑demand supercomputing platform that lets AI teams provision private, single‑tenant GPU clusters with the latest NVIDIA GB300, B200, and H200 accelerators via a web console or API. The service offers up to 64‑GPU nodes with NVLink and InfiniBand interconnects, SOC 2 Type II security, and pay‑as‑you‑go per‑GPU‑hour billing, enabling scalable training and inference for research labs and enterprise ML teams.
Funding: $275M
Rough estimate of the amount of funding raised
JP Morgan
JP Morgan
Funding: $275M
Rough estimate of the amount of funding raised
East New York, United States
Io.net Cloud offers a decentralized computing network that provides machine learning engineers with instant, permissionless access to global GPU resources for their workloads. This platform enables efficient deployment of pre-configured clusters, significantly reducing costs and deployment time for AI startups.
Funding: $35M
Rough estimate of the amount of funding raised
Hack VC
Hack VC
Funding: $35M
Rough estimate of the amount of funding raised
Miami, United States
Hydra Host provides global access to bare metal GPU servers optimized for AI and HPC workloads. The platform aggregates capacity from independent data centers, offering wholesale pricing and eliminating the need for capital expenditure or hardware lock-in. Users gain simplified, unified provisioning via a single API for scalable, high-performance compute resources.
Funding: $14.1M
Rough estimate of the amount of funding raised
Founders Fund
Founders Fund
Funding: $14.1M
Rough estimate of the amount of funding raised
Pasadena, United States
Compute Labs operates an AI infrastructure investment platform that facilitates the financing and securing of GPU supply for cloud and HPC providers. The platform tokenizes physical GPUs into digital assets, enhancing liquidity and providing transparent ownership records. This ecosystem allows investors to gain exposure to the AI compute market through yield-bearing digital assets backed by real hardware.
Funding: $3M
Rough estimate of the amount of funding raised
Protocol Labs
Protocol Labs
Funding: $3M
Rough estimate of the amount of funding raised
Santa Clara, Cuba
EnCharge AI develops high-efficiency analog in-memory computing GPUs and digital AI accelerators for edge-to-cloud deployment. Their validated hardware and flexible software offer significant improvements in performance, TCO, and sustainability compared to traditional solutions. The company provides versatile products from chiplets to PCIe cards, enabling seamless orchestration for on-device and cloud AI inference.
Funding: $44.3M
Rough estimate of the amount of funding raised
DARPA
DARPA
Funding: $44.3M
Rough estimate of the amount of funding raised
Berkeley, United States
ClearML is an AI infrastructure platform that centralizes GPU resource provisioning, model development, and GenAI deployment across on‑premise, cloud, and hybrid environments. Its control plane offers multi‑tenant GPU‑as‑a‑Service with quota management, priority scheduling, and built‑in security, while the integrated IDE provides experiment tracking, data versioning, and automated hyper‑parameter optimization. The platform also includes a GenAI App Engine for rapid LLM and RAG workload serving with role‑based access control and detailed usage billing.
Funding: $13.8M
Rough estimate of the amount of funding raised
Funding: $13.8M
Rough estimate of the amount of funding raised
Sydney, Australia
Provides a multi-cloud AI compute platform that enables real-time GPU resource management, workload migration, and cost optimization across major cloud vendors. By reducing cluster provisioning times to minutes and supporting multi-node training with fixed budgets, it streamlines AI development and inference while maximizing GPU utilization and reducing operational overhead.
Funding: $7.8M
Rough estimate of the amount of funding raised
Blackbird VenturesPeak XV PartnersRebel Fund
Blackbird VenturesPeak XV PartnersRebel Fund
Funding: $7.8M
Rough estimate of the amount of funding raised
XFA AI provides cloud-based GPU rentals through a user-friendly interface, enabling businesses to access high-performance computing resources without the need for significant upfront investment. This service reduces operational costs associated with GPU usage, making advanced computing more accessible for data-intensive applications.
Funding: $8.4M
Rough estimate of the amount of funding raised
Funding: $8.4M
Rough estimate of the amount of funding raised
Delhi, India
Provides a decentralized platform that aggregates idle GPU resources from data centers and independent providers worldwide, creating a scalable and cost-effective infrastructure for on-demand high-performance computing. This system addresses the shortage of AI-grade GPUs by enabling seamless access to thousands of GPUs, including H100s and A6000s, for applications like AI training, rendering, and scientific computation.
Funding: $5.8M
Rough estimate of the amount of funding raised
Funding: $5.8M
Rough estimate of the amount of funding raised
City of New York, United States
Lightning AI provides an AI cloud platform designed for developers and AI teams to efficiently build and deploy high-performance machine learning models. The platform offers specialized tools, collaborative GPU workspaces, managed clusters for training and inference, and pay-per-token APIs. This infrastructure accelerates the entire AI product lifecycle from initial concept to production deployment while offering enterprise-grade security and multi-cloud portability.
Palo Alto, United States
Mithril offers an omnicloud platform that aggregates GPU, CPU, and storage resources across multiple cloud providers for AI workloads. It provides algorithmically determined, demand‑driven pricing and a unified interface with batch SDKs and APIs to enable developers to run asynchronous jobs on cost‑effective spot capacity.
Mountain View, United States
The startup offers an AI training platform that enables instance-less deployment across thousands of GPUs with minimal code, facilitating rapid model training. This technology allows AI engineers to achieve faster training times, improved model performance, and reduced operational costs.
Funding: $80M
Rough estimate of the amount of funding raised
Funding: $80M
Rough estimate of the amount of funding raised
London, United Kingdom
FluidStack provides on-demand access to thousands of NVIDIA A100 and H100 GPUs, enabling AI engineers to rapidly scale their training and inference workloads without long-term contracts. The platform offers fully managed GPU clusters with 24/7 support, significantly reducing operational overhead and accelerating model deployment.
Funding: $4.5M
Rough estimate of the amount of funding raised
Funding: $4.5M
Rough estimate of the amount of funding raised
Calgary, Canada
Denvr Cloud provides on-demand and dedicated GPU computing for AI inference and model training, utilizing NVIDIA GPUs and Intel AI accelerators to enhance performance and scalability. The platform simplifies AI operations by offering transparent pricing and real-time cost monitoring, addressing the need for efficient and cost-effective infrastructure in AI development.
Funding: $10.8M
Rough estimate of the amount of funding raised
Funding: $10.8M
Rough estimate of the amount of funding raised
City of New York, United States
This company provides on-demand GPU cloud infrastructure optimized for AI and machine learning workloads. Their platform offers scalable GPU clusters, high-speed storage, and secure networking, enabling teams to accelerate model training and deployment.
Brisbane, Australia
Bach is a platform-as-a-service that automates the setup and management of scalable cloud hosting environments specifically for AI and GPU workloads, eliminating the need for DevOps expertise. By utilizing multi-tenant cluster sharing and auto-scaling, Bach reduces infrastructure costs and accelerates application development, enabling teams to focus on building rather than managing complex cloud systems.
Palo Alto, United States
Provides a cloud orchestration platform that integrates directly with AWS to deploy and manage AI models and applications on a scalable GPU grid. It eliminates inefficiencies like overprovisioning, idle resources, and reliance on third-party resellers by offering true autoscaling, cost management, and a one-click setup, reducing cloud costs and deployment time significantly.
Funding: $5M
Rough estimate of the amount of funding raised
HOF Capital
HOF Capital
Funding: $5M
Rough estimate of the amount of funding raised
Wichita, United States
The startup provides cloud-based graphics processing services tailored for artificial intelligence research, visual effects production, and data science. By offering scalable computing power for launching AI instances and machine-learning models, it enables organizations to fully utilize their graphics processing capabilities for complex analyses.
Funding: $22M
Rough estimate of the amount of funding raised
Funding: $22M
Rough estimate of the amount of funding raised
Indore, India
HynixCloud provides a unified cloud platform that lets users launch on-demand NVIDIA GPU instances—including H200, H100, A100, L4, and V100—via a web console or API, with integrated compute, storage, and networking. The service offers pay‑as‑you‑go pricing and a free trial, enabling startups, researchers, and enterprises to scale AI training, inference, or graphics workloads without managing physical hardware.
Milpitas, United States
meShare offers a high-performance GPU cloud platform designed for the IoT industry, providing virtualized GPU resources for efficient AI model training and deployment. The platform enables businesses to enhance their operations and product offerings by integrating AI-driven solutions into their processes, significantly reducing costs and improving efficiency.
Funding: $13.5M
Rough estimate of the amount of funding raised
Funding: $13.5M
Rough estimate of the amount of funding raised
San Francisco, United States
Jetify Devspace provides a cloud development environment optimized for building AI applications, featuring GPU-enabled instances and automated setup with Devbox for seamless package management. It allows teams to quickly launch projects with pre-built templates and scale resources efficiently, addressing the need for accessible and powerful infrastructure in AI development.
Funding: $20.5M
Rough estimate of the amount of funding raised
Funding: $20.5M
Rough estimate of the amount of funding raised
Toronto, Canada
Arc Compute provides fully managed, end‑to‑end GPU infrastructure for AI research labs, HPC centers, and enterprise data‑center teams. It offers custom‑designed GPU servers, turnkey AI clusters built on NVIDIA‑validated architectures, and reserved bare‑metal GPU cloud capacity with guaranteed H100 performance, handling everything from planning and procurement to deployment and ongoing optimization. The service enables customers to train, simulate, and infer at scale with high reliability while eliminating hardware complexity.
Open People Network
Swift Compute offers on‑demand dedicated GPU infrastructure featuring NVIDIA A100/H100 accelerators with NVLink, automated provisioning via API/CLI, and pay‑per‑second billing. The platform provides elastic cluster management, real‑time utilization dashboards, and SOC‑2‑compliant security to enable AI teams to scale training and inference workloads efficiently.
Founded 20252K+
San Francisco, United States
Hyperstack is a self-service GPU cloud platform that enables users to deploy between 8 and 16,384 NVIDIA H100 SXM GPUs for high-performance computing, machine learning, and data analytics. The platform addresses the need for scalable, cost-effective GPU resources, allowing businesses to efficiently manage demanding workloads without the constraints of traditional cloud providers.
AMSTERDAM, Netherlands
HyperAI provides HyperCLOUD, a European‑focused cloud platform that gives businesses and developers on‑demand access to Nvidia A100 and H100 GPUs for AI workloads. The service offers a self‑service portal, pay‑as‑you‑go pricing, and full data ownership with compliance to regional privacy regulations, enabling enterprises, startups, and research teams to run high‑performance AI models without vendor lock‑in.
Boston, United States
Dash Cloud offers a GPU cloud infrastructure that provides access to NVIDIA graphics processing units for applications in machine learning, rendering, and cloud gaming at a cost reduction of up to 80%. The platform ensures rapid provisioning of services within one business day, backed by a 99.95% uptime guarantee and robust physical security measures.
London, United Kingdom
Oblivus offers a cloud platform that provides on-demand and reserved access to a range of NVIDIA GPUs and modern CPUs with per‑minute billing and no quota limits. Users can launch, modify, and scale instances instantly via a fast API or dashboard, benefiting from pre‑installed drivers, deep‑learning libraries, and high‑speed networking at prices up to 80% lower than major providers.
Taipei, Taiwan
Neurowatt AI operates as an Applied AI Foundry, delivering AI solutions through a full-stack infrastructure approach. The company offers a hybrid edge cloud computing platform for GPU rental and on-premise modular data centers for secure AI deployment. They provide custom AI agents and services designed to transform compute power into strategic assets for enterprise digital transformation.
San Diego
1Legion provides dedicated bare‑metal GPU servers and scalable clusters with full access to NVIDIA H100/H200 accelerators, delivering consistent performance and predictable pricing for AI, ML, media rendering, and HPC workloads. The platform offers unmetered bandwidth, pre‑installed AI frameworks or custom OS images, secure private networking, and API‑driven integration for seamless scaling from single servers to thousands of GPUs.
San Diego
1Legion provides dedicated bare‑metal GPU servers and scalable clusters with full access to NVIDIA H100/H200 accelerators, delivering consistent performance and predictable pricing for AI, ML, media rendering, and HPC workloads. The platform offers unmetered bandwidth, pre‑installed AI frameworks or custom OS images, secure private networking, and API‑driven integration for seamless scaling from single servers to thousands of GPUs.
TensorDock is a marketplace that provides access to high-performance GPU and CPU cloud services optimized for AI, deep learning, and rendering. The platform offers cost-effective compute resources, enabling users to scale their workloads efficiently.
Funding: $100K
Rough estimate of the amount of funding raised
Funding: $100K
Rough estimate of the amount of funding raised
East New York, United States
Build AI provides a cost-effective GPU cloud platform for training machine learning models by utilizing interruptible workloads powered by renewable energy, allowing users to save up to 50% on training costs. The service addresses the high energy expenses associated with AI model training by strategically pausing operations during peak hours, enabling more efficient use of resources.
Founded 2024100+