Tutorials
Step-by-step guides for GPU workloads and platform features.
Topics
Jupyter Notebooks
Deploy and run Jupyter Notebooks on GPU virtual machines.
Docker
Use Docker containers on Hyperstack virtual machines.
Ollama
Run open-source LLMs on a GPU virtual machine with a REST API and OpenAI-compatible endpoint.
Open WebUI
Add a browser-based chat interface to an Ollama model running on a GPU virtual machine.
vLLM
Serve open-source LLMs behind an OpenAI-compatible API on a single GPU virtual machine.
Multi-GPU vLLM
Serve a large model across multiple GPUs on one virtual machine with vLLM tensor parallelism.
LLM on Kubernetes
Serve a large language model across a multi-node Kubernetes cluster with vLLM.
Fine-Tuning LLMs with Unsloth
Fine-tune Qwen3-8B with QLoRA on a GPU virtual machine using Unsloth.
Ephemeral Drive Mounting
Customize ephemeral drive mounting behavior using cloud-init.
Was this page helpful?