AI product Open source · MIT
stable-diffusion.cpp is a pure C/C++ inference implementation for diffusion image and video models, based on ggml and designed to work similarly to llama.cpp. It provides a command-line executable for generating and editing images from text or image inputs, with support for model families including Stable Diffusion, SDXL, SD3, FLUX, Qwen Image, Wan, LTX, and other listed image and video models. The project supports PyTorch checkpoints, Safetensors, and GGUF weights, and can convert weights to GGUF or Safetensors. It runs on CPU, CUDA, Vulkan, Metal, OpenCL, and SYCL backends across Linux, macOS, Windows, and Android via Termux; features include LoRA, ControlNet, IP-Adapter, latent-consistency models, VAE tiling, ESRGAN upscaling, quantization, and an embedded web UI. The repository is under active development, so its API and command-line options may change frequently.