Skip to main content

Tutorials

Step-by-step guides for GPU workloads and platform features.

Jupyter Notebooks

Deploy and run Jupyter Notebooks on GPU virtual machines.

Docker

Use Docker containers on Hyperstack virtual machines.

Ollama

Run open-source LLMs on a GPU virtual machine with a REST API and OpenAI-compatible endpoint.

Open WebUI

Add a browser-based chat interface to an Ollama model running on a GPU virtual machine.

vLLM

Serve open-source LLMs behind an OpenAI-compatible API on a single GPU virtual machine.

Multi-GPU vLLM

Serve a large model across multiple GPUs on one virtual machine with vLLM tensor parallelism.

LLM on Kubernetes

Serve a large language model across a multi-node Kubernetes cluster with vLLM.

Fine-Tuning LLMs with Unsloth

Fine-tune Qwen3-8B with QLoRA on a GPU virtual machine using Unsloth.

Ephemeral Drive Mounting

Customize ephemeral drive mounting behavior using cloud-init.