Qwen3.6-35B-A3B-EXTENSOR
An EXTENSOR runtime image derived from Qwen/Qwen3.6-35B-A3B. It contains the text target and its MTP draft module; the...
An EXTENSOR runtime image derived from Qwen/Qwen3.6-35B-A3B. It contains the text target and its MTP draft module; the...
Calibrated 5-bit MLX quantization of poolside/Laguna-S-2.1 (118B total, 8B activated per token), produced with oMLX oQ at level...
Calibrated 3-bit MLX quantization of poolside/Laguna-S-2.1 (118B total, 8B activated per token), produced with oMLX oQ at level...
MLX 4-bit affine quantization of poolside/Laguna-S-2.1, packaged for Apple Silicon experiments and local OpenAI-compatible serving. Laguna S 2.1...
This repository provides an INT8-ConvRot quantized version of the Qwen3-4B model. It has been specifically tailored for use...
Inkling, self-quantized to GGUF by Atomic Chat. Built straight from Thinking Machines Lab’s original weights with a per-tensor...
> [!Note] > This repository contains model weights and configuration files for Agents-A1 in the Hugging Face Transformers...
Модель генерации текста Модальности:Генерация текста Области применения:Диалог / чат Задача: Генерация текста Автор: katafiek Теги: gguf, roleplay, uncensored,...
High-quality imatrix GGUF quantizations of tencent/Hy3, Tencent’s 295B-parameter Mixture-of-Experts model with ~21B active parameters per token. Produced with...
Массы квантуют до INT4 (размер группы 128); активации выполняют в BF16. Результатом является контрольная точка объемом 388 ГБ...