Tiamat-8b-1.2-Llama-3-DPO-GGUF
Оригинальная модель: https://huggingface.co/Gryphe/Tiamat-8b-1.2-Llama-3-DPO Все кванты сделаны с использованием опции imatrix с набором данных, предоставленным Kalomaze здесь Если модель...
Оригинальная модель: https://huggingface.co/Gryphe/Tiamat-8b-1.2-Llama-3-DPO Все кванты сделаны с использованием опции imatrix с набором данных, предоставленным Kalomaze здесь Если модель...
Оригинальная модель: https://huggingface.co/cognitivecomputations/dolphin-2.9-llama3-8b-1m Все кванты сделаны с использованием опции imatrix с набором данных, предоставленным Kalomaze здесь Если модель...
— NOT Updated for new pre-tokenizer fixes (yet), I recommend using bartowski’s quants. https://huggingface.co/bartowski/Meta-Llama-3-70B-Instruct-GGUF — quants done with...
Imatrix-guided APEX quantization of IFM/K2-Horizon-MoVA-36B-A4B — MBZUAI’s Institute of Foundation Models (the LLM360/K2 lineage), released 2026-09-01, Apache-2.0. Not...
Apodex-1.1 is a reasoning-first model for complex, long-horizon research tasks. Beyond searching and writing reports, it works directly...
Two quantizations of the DFlash 2 draft model for Qwen/Qwen3.8-27B, both calibrated with an importance matrix captured from...
HuggingFace’s download widget does not recognize KS and KT ggufs ——> Text-only GGUF conversions of Qwen3.8-27B, quantized with...
Community GGUF quantizations of nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16. These files contain the standard 52-layer inference model. The trailing MTP prediction head...
Inkling, self-quantized to GGUF by Atomic Chat. Built straight from Thinking Machines Lab’s original weights with a per-tensor...
«This is humanity’s race. The solution is open source. Stay sovereign.» — AIOpsInSpace 📑 Table of Contents 🎯...