On-demand NVIDIA GPU instances
Provision single-instance GPU capacity in under a minute, with selectable configurations from 1x to 8x GPUs.
Verda is a full-stack AI cloud for running AI workloads on GPU infrastructure. It provides on-demand NVIDIA GPU instances and self-service multi-node clusters with managed Slurm or Kubernetes.
Verda is a full-stack AI cloud focused on infrastructure for AI workloads. It combines on-demand NVIDIA GPU instances with self-service, multi-node clusters and managed Slurm or Kubernetes options. Users can choose individual GPU capacity for short-lived or focused jobs, or provision a multi-node environment when an application requires distributed compute.
Provision single-instance GPU capacity in under a minute, with selectable configurations from 1x to 8x GPUs.
The listed instance hardware includes NVIDIA B300, B200, H200, H100, and A100 GPUs, allowing users to select from the models exposed by the platform.
Create clusters without a manual setup process or a sales call, supporting workloads that need more than one node.
Choose between managed Slurm and Kubernetes for cluster operation, depending on the workload or orchestration workflow.
GPU instances are offered with pay-as-you-go pricing; the available evidence does not state the applicable rates.
Run inference workloads that need several GPUs, such as the multi-GPU inference described in Verda’s 1XWM generative video collaboration.
Supply GPU infrastructure for video-generation systems where higher video quality is part of the application workflow.
Use an instant multi-node cluster when an AI workload must run across multiple machines rather than a single GPU instance.
Provision a cluster with Slurm or Kubernetes when a team needs an established scheduler or container-orchestration environment without setting up the cluster manually.
The GPU instances page states that on-demand instances can be provisioned in under a minute.
Instant clusters are described as provisioning in 20 minutes.
Verda offers managed Slurm or Kubernetes for its instant clusters.
The source lists NVIDIA B300, B200, H200, H100, and A100 options, among others.
The source identifies pay-as-you-go pricing for GPU instances, but the available material does not include specific rates or plan limits.
salad.com
面向 AI 工作负载的分布式 GPU 云,按用量计费
www.hyperstack.cloud
Hyperstack is a cloud GPU platform for running AI and machine learning workloads, including training, inference, data analytics, and model development. It also provides AI Studio, virtual machines, and managed Kubernetes for deploying and operating GPU-backed workloads.
radiant.co
Radiant is an integrated AI infrastructure platform that finances, builds, and operates data centers, GPU systems, networking, storage, and managed services. It helps AI teams and infrastructure operators provision and run compute through a unified platform and FlightDeck control plane.
together.ai
Together AI 是一个支持推理、微调、GPU 集群、沙盒和托管存储的 AI 云平台。
www.paperspace.com
Paperspace is a cloud platform for developing, training, and deploying machine learning applications with managed notebooks, GPU machines, and deployment workflows. It serves ML developers, data scientists, researchers, and teams that need on-demand accelerated computing.
deepinfra.com
DeepInfra provides hosted machine-learning model inference and on-demand GPU instances for developers and teams. Its catalog covers text, image, audio, video, embedding, reranking, and other model workloads with pay-as-you-go pricing.