Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
-
Updated
Sep 9, 2026 - Go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Android native AI inference library, bringing text, image, video, STT, TTS inference
A UI for stable-diffusion.cpp.
A thin cython wrapper around llama.cpp, whisper.cpp and stable-diffusion.cpp
single-executable / library which combines llama.cpp, whisper.cpp, and stable-diffusion.cpp
Examples using the llmedge library
An experimental nanobind wrapper around llama.cpp, whisper.cpp, and stable-diffusion.cpp
python package to build a ggml/llama.cpp/whisper.cpp/stable-diffusion.cpp stack
A tiny C++17 / ggml image generator for FLUX.2 [klein] by Black Forest Labs, built on top of stable-diffusion.cpp. No Python, no PyTorch at build or run time, runs on both GPU, or CPU-only targets
FLAI is a self-hosted, privacy-first AI platform. Local assistant for chat, voice, image/video gen, doc Q&A & camera analysis. Open source, GPU-optimized, multi-user with request queuing. Data never leaves your machine.
Pre-built stable-diffusion.cpp binaries for Leaxer
Seven local-inference studies on one 32 GB Apple Silicon machine: what ships, what it costs, and how it fails.
Espresso! - AI image generation made simple.
Image Generation module for I4.0
Universal Edge AI ASIC: Open-Source Silicon Accelerator for LLMs (llama.cpp GGUF Q4_K/IQ) & Diffusion DiT (Flux, SDXL) in SkyWater SKY130 130nm CMOS.
Cutting local AI server build cost with mixed used GPUs - measured data, reproduction scripts, and engine patches (KO/EN)
llama.cpp for coding agents + stable-diffusion.cpp for image generation
Portable native inference stack for Mage-Flow-Turbo using stable-diffusion.cpp/sd-cli, Q8_0 DiT GGUF, Qwen3-VL-4B Q4_K_M and a dedicated VAE, with CPU/CUDA backends, CLI, REST API, model verification and Kaggle integration.
To associate your repository with the stable-diffusion-cpp topic, visit your repo's landing page and select "manage topics."