A coding agent connected to a local Ollama model
Use Ollama on your hardware platform to run a local coding model and connect a CLI coding agent. This playbook uses Claude Code with Ollama's built-in launch method (ollama launch claude), so you can work without environment variables, provider config files, or external cloud APIs.
This playbook uses Claude Code as the CLI agent, connected to a local Ollama model for inference.
You'll run a local coding model (Qwen3.6 qwen3.6:27b) on your hardware platform with Ollama, launch Claude Code against it with a single command, and complete a small coding task end-to-end.
Required:
Optional:
Use the matrix below to confirm your hardware platform, recommended default local settings, and whether multi-node applies.
| Hardware platform | OS | Memory | Recommended default local settings | Multi-node capable hardware |
|---|---|---|---|---|
| DGX Station | DGX OS (Linux) | Large HBM + Grace DRAM | Ollama + qwen3.6:27b; Claude Code via ollama launch | — |
Hardware requirements
qwen3.6:27bSoftware requirements
ollama launch and the recommended model): ollama --versionvenv support for the coding-task verificationBrowse models you can pull and run with Ollama in the Ollama library. Use the tags and sizes that fit your hardware platform’s memory and storage.
| Hardware platform | More recipes |
|---|---|
| DGX Station | Ollama library · Qwen3.6 |
Use the Claude Code tab for the base workflow.
No local ancillary files are required. All steps use Ollama and Claude Code on your hardware platform.
ollama launch and the recommended model tag — verify with ollama --version and install from ollama.com/download if needed~/.ollama/models (see Cleanup in the Claude Code tab). Cleanup is optional and removes downloaded model files.ollama launch on supported hardware platforms