AI & ML Local Development Quick-Start
TunaOS gives AI and ML developers a repeatable desktop base for local language models, vision models, and container training pipelines.
Why Container-Native for AI?
- Zero Host Pollution: The drivers and the base libraries stay as they are. Your CUDA, ROCm, and PyTorch stacks stay inside OCI containers or Flatpaks.
- Repeatable Toolchains: Share the same container dev environment (a
Containerfileor a devcontainer) with the other people on your team. - Podman and OCI Native: Podman comes with the image. Use it to run an inference server such as Ollama, LocalAI, or vLLM.
Quick Setup
1. Local LLM Runner (Ollama via Podman)
Run a local model with GPU passthrough. The host packages do not change:
podman run -d --device nvidia.com/gpu=all -v ollama:/root/.ollama -p 11434:11434 --name ollama ollama/ollama
The nvidia.com/gpu device needs the NVIDIA container toolkit on the host. Generate the CDI specification first. Then check that the device name resolves before you start the container.
2. Podman Desktop
Install Podman Desktop as a Flatpak to see and control your local containers and models:
flatpak install flathub io.podman_desktop.PodmanDesktop