Install
Get the app
curl -fsSL https://raw.githubusercontent.com/kyndlo/tapioca/main/scripts/install.sh | shmacOS Apple Silicon. Windows users can grab the x64 or ARM64 bundle from Releases.
Local AI, minus the drama
Run language, image, video, and coding-agent models on your own Mac or Windows PC—with one small, friendly command-line tool.

› tapioca run qwen3:8b-q4_k_mmodel ready · 6.2 GB · metalyou › Build me something delightful.tapioca › Let's cook. ▋01 · Beginner lane
No model archaeology degree required. Tapioca downloads what you need and remembers where it lives.
Install
curl -fsSL https://raw.githubusercontent.com/kyndlo/tapioca/main/scripts/install.sh | shmacOS Apple Silicon. Windows users can grab the x64 or ARM64 bundle from Releases.
Choose
tapioca catalogSee download size, memory guidance, GPU needs, and the right variant for your computer.
Run
tapioca run qwen3:8b-q4_k_mIf the model is missing, Tapioca pulls it automatically. Type /bye when you're done.
Not sure which model? Use the friendly model chooser →
02 · Bring your computer
Metal, MLX, MFLUX, and native llama.cpp. A particularly cozy home for local models.
Vulkan text models, NVIDIA CUDA, plus DirectML image generation for AMD and Intel.
Native ARM64 Tapioca, CPU llama.cpp, and ONNX image generation without x64 emulation.
03 · Six verbs, lots of power
Learn the shape once, then move from a chat to an API, an image, or a full coding agent.
Download a model once
tapioca pull qwen3:8b-q4_k_mChat in your terminal
tapioca run qwen3:8b-q4_k_mOpen an API for your apps
tapioca serve qwen3:8b-q4_k_mCreate images locally
tapioca image sd-turbo --prompt "A pearl astronaut"Turn prompts into motion
tapioca video ltx-video:2b-fp16 --prompt "Clouds rolling in"Power your coding agent
tapioca launch opencode qwen3-coder:30b-mlx04 · The local studio
Tapioca understands platform-native diffusion: MLX and MFLUX on Mac, CUDA and DirectML on Windows, and native ONNX on Windows ARM64.
IMAGE RECIPEfox-studio
base sd-turbo:onnx-directmlprompt “a red fox in snow”size 512 × 512steps 4tapioca image fox-studio05 · Expert lane
Use Tapioca as an orchestration layer, a local OpenAI-compatible endpoint, or a portable runtime for your team.
Tapiocacatalog · routing · lifecycleUse /v1/chat/completions with tools, streaming, and familiar clients.
A model variant declares its runtime, memory guidance, GPU needs, and platform.
Save reusable recipes that combine a base model, LoRAs, adapters, and presets.
Models and generated media stay on your machine under ~/.tapioca.

Your machine has been waiting
Start small. Pull a model. Make something weird and wonderful.