This setup runs Ollama in Docker on a Linux host with:
- NVIDIA GPU acceleration
- Ollama accessible from VS Code/Continue at
127.0.0.1:11434 - No inbound LAN access to Ollama
- No outbound internet/LAN access from the Ollama container
- Persistent model storage
- A temporary internet-enabled container for downloading models