diffusiongemma-26B-A4B-asym-2bitexp-GGUF
An antirez-style asymmetric low-bit GGUF quant of diffusiongemma-26B-A4B-it, a Gemma-4 MoE diffusion language model (26B total parameters, ~4B...
An antirez-style asymmetric low-bit GGUF quant of diffusiongemma-26B-A4B-it, a Gemma-4 MoE diffusion language model (26B total parameters, ~4B...
> A reasoning critic — evaluates and improves chain-of-thought. Base model: fableforge-ai/ReasonCritic-7B · Part of the FableForge ecosystem....
MagicQuant is a benchmark driven GGUF hybrid discovery and validation system focused on finding real, practical GGUF quants...
Original Model: OrionLLM/GRM-2.6-Opus Architecture: Qwen3.6-27B License: Apache 2.0 MTP Support: Yes Prompt: Create a complete SVG loading animation...
MagicQuant is a benchmark driven GGUF hybrid discovery and validation system focused on finding real, practical GGUF quants...
Cerebellum v5 is an ablation-guided mixed-precision GGUF quantization of google/gemma-4-26B-A4B-it. This is a 26B-parameter MoE model with 4B...
— llama.cpp — ramalama — LM Studio — koboldcpp — Jan AI — Text Generation Web UI —...
— Used mmangkad/Qwen3.6-27B-NVFP4 as base model (Thank you for doing ModelOpt quant!) — Another mixed precision quant: ssmout...
This is a high-performance, hybrid Mixture of Experts (MoE) GGUF version of Google’s Gemma 4 26B. By utilizing...
Original model: https://huggingface.co/allura-org/Qwen3.5-27B-Anko — llama.cpp — ramalama — LM Studio — koboldcpp — Jan AI — Text Generation...