A self-hosted browser interface with models running locally on your GPU
TIP
Use this tab when you are working on a local desktop session on the hardware platform, or when you want to run Docker commands directly. For remote laptop access through NVIDIA Sync, use the Open WebUI Remotely tab instead.
You should first make sure you can run Docker commands without entering your sudo password.
To test that, open a terminal on the hardware platform and run:
docker ps > /dev/null
Success: If the command returns a blank, then skip ahead to Step 2.
Otherwise, you will see a permission denied error, which means you still need to remove the sudo requirement. To do that, add your user to the docker group with the commands below.
sudo usermod -aG docker $USER
newgrp docker
Then verify the change is set by testing Docker access again with the command:
docker ps > /dev/null
Pull the container image onto your hardware platform with the command:
docker pull ghcr.io/open-webui/open-webui:ollama
Wait for the image to download, then go to Step 3.
Start the Open WebUI container by running:
docker run -d -p 8080:8080 --gpus=all \
-v open-webui:/app/backend/data \
-v open-webui-ollama:/root/.ollama \
--name open-webui ghcr.io/open-webui/open-webui:ollama
This will start the Open WebUI container and make it accessible at http://localhost:8080. You can access the Open WebUI interface from your local web browser.
NOTE
Application data will be stored in the open-webui volume and model data will be stored in the open-webui-ollama volume.
Set up the initial administrator account for Open WebUI. This is a local account that you will use to access the Open WebUI interface.
You'll then download a language model through Ollama and configure it for use in Open WebUI. This download happens on your hardware platform and may take several minutes.
gpt-oss:20b in the search field.You can verify that the setup is working properly by testing model inference through the web interface.
Try downloading different models from the Ollama library at https://ollama.com/library.
For agentic workloads, see the Agent-ready Models tab for the recommended model on your hardware platform.
You can also use the Open WebUI Remotely tab so that you can reach the same setup from your laptop through NVIDIA Sync and monitor GPU and memory usage as you try different models.
If Open WebUI reports an update is available, you can update the container image by running:
docker pull ghcr.io/open-webui/open-webui:ollama
Use these steps when you want to completely remove the Open WebUI installation and free up resources. Cleanup is optional rollback—not required to finish the playbook.
WARNING
These commands permanently delete all Open WebUI data and downloaded models.
Stop and remove the Open WebUI container:
docker stop open-webui
docker rm open-webui
Remove the downloaded images:
docker rmi ghcr.io/open-webui/open-webui:ollama
Remove persistent data volumes:
docker volume rm open-webui open-webui-ollama