diffusiongemma
Here are 9 public repositories matching this topic...
FastMCP fleet MCP server for diffusion LMs (dLLM). DiffusionGemma on Goliath RTX 4090 — batch inference, HLE-shaped reasoning, ~200–400 tok/s. Doc phase; llama-diffusion-cli sidecar next. Complements local-llm-mcp.
-
Updated
Jun 17, 2026
DiffusionGemma node pack for ComfyUI, built on MCP backbone for fully-agentic consumption — discrete diffusion text generation with per-step canvas snapshots, commit heatmaps, and structured trace data. Watch meaning crystallize out of noise.
-
Updated
Aug 14, 2026 - Python
Docker-Compose template to self-host Google DiffusionGemma 26B on an NVIDIA GPU host via llama.cpp
-
Updated
Jun 12, 2026 - Dockerfile
Matrix-style logit conditioning for DiffusionGemma's llama.cpp denoiser
-
Updated
Jun 15, 2026 - Python
Native WinUI 3 control panel for running local llama.cpp and DiffusionGemma backends with model management, Hugging Face downloads, runtime tuning, logs, and resource monitoring.
-
Updated
Jun 11, 2026 - C#
llama.cpp fork with experimental DiffusionGemma full-GPU and CUDA fusion support
-
Updated
Jul 16, 2026 - C++
Local DiffusionGemma coding agent for Windows, WSL2, and 16 GB NVIDIA GPUs
-
Updated
Jul 31, 2026 - Python
Add this topic to your repo
To associate your repository with the diffusiongemma topic, visit your repo's landing page and select "manage topics."