A thin cython wrapper around llama.cpp, whisper.cpp and stable-diffusion.cpp
-
Updated
Oct 9, 2026 - Python
A thin cython wrapper around llama.cpp, whisper.cpp and stable-diffusion.cpp
An experimental nanobind wrapper around llama.cpp, whisper.cpp, and stable-diffusion.cpp
python package to build a ggml/llama.cpp/whisper.cpp/stable-diffusion.cpp stack
FLAI is a self-hosted, privacy-first AI platform. Your all-in-one local assistant: smart chat, web search, voice, media generation, and document analysis. Offers long-term memory, Model Hub, and OpenAI API. Fully optimized for GPU and CPU environments.
Plug-and-play Stable Diffusion engine (sd-cli first: CUDA/ROCm/Vulkan/CPU) with Gradio WebUI
Reproducible Kaggle NVIDIA T4 x2 workflow for Mage-Flow-Turbo and Mage-Flow-Edit-Turbo, with authenticated REST API, dual-GPU CUDA routing, deterministic inference evidence, bilingual documentation, and exact-revision release validation.
Seven local-inference studies on one 32 GB Apple Silicon machine: what ships, what it costs, and how it fails.
Bulk AI image generation on your own Windows PC: FLUX.2, Qwen-Image, SDXL and more via stable-diffusion.cpp. Batches, reference images, LoRAs, upscaling, an OpenAI-compatible API and an MCP server for AI agents. No cloud, no account.
Espresso! - AI image generation made simple.
Image Generation module for I4.0
Universal Edge AI ASIC: Open-Source Silicon Accelerator for LLMs (llama.cpp GGUF Q4_K/IQ) & Diffusion DiT (Flux, SDXL) in SkyWater SKY130 130nm CMOS.
Cutting local AI server build cost with mixed used GPUs - measured data, reproduction scripts, and engine patches (KO/EN)
Reproduction: orchestrating educational yonkoma (4-panel manga) production with Stable Diffusion on Vast.ai
A local, chat-first studio for Qwen-Image-2.1: generation and editing on your own PC with CUDA, Vulkan or CPU-only backends, a native save dialog and in-app guides. AI-generated under human supervision (unofficial).
Portable native inference stack for Mage-Flow-Turbo using stable-diffusion.cpp/sd-cli, Q8_0 DiT GGUF, Qwen3-VL-4B Q4_K_M and a dedicated VAE, with CPU/CUDA backends, CLI, REST API, model verification and Kaggle integration.
To associate your repository with the stable-diffusion-cpp topic, visit your repo's landing page and select "manage topics."