Ollama
Large Language Model API
Get up and running with Llama 3.2, Mistral, Gemma 2, and other large language models.
Links:
- Home: Ollama
- Source: GitHub - ollama/ollama: Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
GPU override for Ollama โ AMD iGPU inference via the Vulkan backend. Enable by setting GPU_COMPOSE_SUFFIX=amdgpu in config/docker//.env NOTE: deliberately the STANDARD image, not :rocm โ colony’s Vega iGPU (Ryzen 7 5825U, gfx90c) has no rocblas kernels in ROCm builds (“dropping ROCm device”), while the standard image ships the ggml Vulkan backend (the :rocm image does not), which supports Vega iGPUs via RADV. OLLAMA_IGPU_ENABLE opts the iGPU into scheduling on Linux. References:
- Vulkan backend for AMD/Intel GPUs: Add Vulkan GPU Backend for AMD/Intel Support ยท Issue #11247 ยท ollama/ollama
- Hardware support matrix (Vulkan path for iGPUs): Hardware support - Ollama