saturn-31b-Q4_K_M-GGUF
A fine-tune of Gemma 4 31B (dense, instruct), trained on Fireworks AI and merged + quantized to Q4KM...
A fine-tune of Gemma 4 31B (dense, instruct), trained on Fireworks AI and merged + quantized to Q4KM...
A LoRA fine-tune of Gemma 4 12B trained on syntetic multi-turn conversational data from the visual novel My...
This is a roleplay and creative-writing merge of google/gemma-4-31B-it. I built it because the usual merge recipe was...
> No matter your GPU. No matter your RAM. With ~4.5 GB of VRAM or unified memory free,...
Mixed-precision MLX quantization of huihui-ai/Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated, quantized with MLX Smart Quantize (MSQ) — my own sensitivity-based mixed-precision quantization method...
A web-design generator fine-tuned from google/gemma-4-31B-it. Give it a natural-language brief; it returns a self-contained front-end as three...
> VIDRAFT attention + Qwen3-4B / Gemma4-E4B FFN crossbreed. > A Qwen3-4B × Gemma4-E4B hybrid — NOT from-scratch....
Added a Jinja chat template so the model can format conversations correctly and work smoothly with mlx-lm chat-style...
> Built with mlx-optiq, the MLX-native toolkit to quantize, fine-tune, and serve LLMs locally on Apple Silicon, no...
> TL;DR — A local Python-coding assistant that thinks before it codes. 8.25 GB, runs on one 16...