Skip to content
#

gguf-quantization

Here are 30 public repositories matching this topic...

A 501M-parameter language model trained from scratch on 20B tokens: custom BPE tokenizer, original Llama-style architecture, training and SFT pipeline, and unattended cloud orchestration. 54 hours on one H100 for about $165.

  • Updated Aug 13, 2026
  • Python

Run Qwen-Image 2.1 locally on a 16 GB Apple Silicon Mac. Free, offline AI image editor and generator with mask inpainting, 4-step Turbo LoRA, 4-bit GGUF models, upscaling and a private mode. Built on stable-diffusion.cpp.

  • Updated Oct 2, 2026
  • HTML

Add this topic to your repo

To associate your repository with the gguf-quantization topic, visit your repo's landing page and select "manage topics."

Learn more