CLI/GUI tool for efficient and easy safetensors and gguf model conversion
-
Updated
Oct 6, 2026 - Zig
CLI/GUI tool for efficient and easy safetensors and gguf model conversion
GGUF Quantization support for native ComfyUI models
Evolution process to find the best quant tensor weights to build the most optimal GGUF options for an AI model.
Análise Avançada de Dados com Causalidade e Aprendizado por Reforço
Fatrocu converts your e-archive invoices, cash register receipts, and receipts into Excel format within minutes using a 100% locally running AI, drastically reducing manual data entry.
Which language models will actually run well on your machine. Measured, not looked up. Free, offline, for Windows, macOS and Linux.
Auto GGUF Converter for HuggingFace Hub Models with Multiple Quantizations (GGUF Format)
Gemma-4-It fine-tuned on PubMedQA using SFT & RLVR
Convert and quantize llm models
Run large language models locally on Windows. A native Rust desktop app and local LLM API server for GGUF models fully offline, no Python, no Docker, no cloud account. It measures your GPU and the model file, then loads a configuration that fits. Tuned on an 8 GB laptop card.
Convert Safetensor weights to GGUF format using google colab and HuggingFace. Use the power of AI locally :)
Power up your projects with Glyvex: All-in-one open-source AI suite for automation and innovation.
High-performance LLM compression engine using SVD matrix decomposition, INT8 quantization, and output caching training acceleration.
Convert Hugging Face models to GGUF with xet support.
A 501M-parameter language model trained from scratch on 20B tokens: custom BPE tokenizer, original Llama-style architecture, training and SFT pipeline, and unattended cloud orchestration. 54 hours on one H100 for about $165.
AI Toolchain in Pure Zig, No Python, No C++
Quantize LLMs automatically.
Run Qwen-Image 2.1 locally on a 16 GB Apple Silicon Mac. Free, offline AI image editor and generator with mask inpainting, 4-step Turbo LoRA, 4-bit GGUF models, upscaling and a private mode. Built on stable-diffusion.cpp.
Autonomous, Sovereign & Zero-Cloud Local Generative AI Workstation
Offline LLM chat for Android & iOS. Download a GGUF model once, then chat fully on-device. No account, no cloud. Made possible with llama.rn and React Native.
To associate your repository with the gguf-quantization topic, visit your repo's landing page and select "manage topics."