[ai]
$ cloudrig deploy — workload ai

Train and run any model. On a real GPU.

From 16GB to 80GB VRAM available. ComfyUI, Automatic1111, LM Studio, Ollama — root access, CUDA pre-installed, no shared notebooks.

supported_tools

Full admin access means you can install anything. Here are the most popular stacks our users run.

ComfyUI

Node-based Stable Diffusion workflows with full GPU acceleration.

Automatic1111

The most popular Stable Diffusion interface, pre-configured for NVIDIA.

LM Studio

Run local LLMs — Llama, Mistral, Mixtral, and more.

Ollama

CLI LLM runner. Pull and run models in seconds.

Fooocus

Simplified Stable Diffusion — great for fast iteration.

KoboldCpp

Run GGUF models with GPU offloading for text generation.

InvokeAI

Professional Stable Diffusion toolkit with a polished UI.

LoRA training

Fine-tune diffusion models on your own data.

vram_matrix

Choose the right tier based on what you want to run.

modelvram neededstarter 16GBstandard 24GBpro 48GBpower 80GB
Stable Diffusion XL8–12 GByesyesyesyes
Flux.1 (dev/schnell)12–24 GBtightyesyesyes
Llama 3 8B (Q4)~6 GByesyesyesyes
Llama 3 70B (Q4)~40 GBnonotightyes
Mixtral 8x7B (Q4)~26 GBnotightyesyes
SDXL LoRA training16–24 GBtightyesyesyes
recommended_tiers

Standard

gpu
RTX 4090
vram
24 GB
$0.60/hr

great for SD and smaller LLMs

recommended

Pro

gpu
L40S
vram
48 GB
$0.90/hr

more VRAM for bigger models

Power

gpu
A100 80GB
vram
80 GB
$1.20/hr

run 70B+ models and big batches

view all plans →

$ run your first model — $0.60/hr