Quick start guide
- Prerequisites
- Starting a virtual machine with cudoctl
- Installing Ollama via SSH
- Using Docker to start a LLM API
Prerequisites
- Create a project and add an SSH key
- Download the CLI tool
Starting a virtual machine with cudoctl
Start a virtual machine with the base image you require, here we will start with an image that already has NVIDIA drivers. You can use the CUDO Compute console to start a virtual machine using the Ubuntu 22.04 + NVIDIA drivers + Docker image or alternatively use the command line toolcudoctl
To use the command line tool you will need to get an API key from the CUDO Compute console, see here: API key
Then run cudoctl init and enter your API key.
First we search to find a virtual machine type to start
epyc-milan-rtx-a4000 (16GB GPU) in the se-smedjebacken-1 data center and image ubuntu-2204-nvidia-535-docker-v20240214 we can start a virtual machine:
Installing Ollama via SSH
Get the IP address of the virtual machineUsing Docker to start a LLM API
If you had created a vm in the previous step delete it by running:start-ollama.txt
-start-script-file start-ollama.txt:
gemma:7b