Qwen3-Coder-Next-NVFP4-GB10
Весовые коэффициенты идентичны байтам вышестоящего количества (проверено config.json и model.safetensors.index.json SHA-256). Это зеркало существует для обеспечения закрепленной, стабильной,...
Весовые коэффициенты идентичны байтам вышестоящего количества (проверено config.json и model.safetensors.index.json SHA-256). Это зеркало существует для обеспечения закрепленной, стабильной,...
— Base model: deepseek-ai/DeepSeek-V4-Flash — Source revision: 6e763230a9d263eca2023f1d4a5ce1bfe126cf48 — Architecture: DeepseekV4ForCausalLM — Model type: deepseekv4` — Tooling branch:...
Partial GPTQ Int4 quant of Qwen/Qwen3.6-27B, produced with the verbatim recipe from Qwen’s own Qwen/Qwen3.5-27B-GPTQ-Int4 — only MLP/FFN...
— Used mmangkad/Qwen3.6-27B-NVFP4 as base model (Thank you for doing ModelOpt quant!) — Another mixed precision quant: ssmout...
Data-driven mixed-precision native TurboQuant checkpoint of Qwen/Qwen3.6-35B-A3B. Extends -TQ-apex2 by skipping the shared-expert down-projection — the tensor family...
GGUF quantizations of DJLougen/Ornstein3.6-35B-A3B-SABER for use with llama.cpp, ollama, LM Studio, and compatible runtimes. Source model is the...
6-bit base mixed-precision quantization of Qwen/Qwen3.6-35B-A3B for Apple Silicon, using the Unsloth Dynamic quantization strategy via mlx-node. Benchmarked...
3-bit base mixed-precision quantization of Qwen/Qwen3.6-35B-A3B for Apple Silicon, using the Unsloth Dynamic quantization strategy via mlx-node. Benchmarked...
— Quantization: oQ6 (sensitivity-driven mixed precision, groupsize=64) — Format: MLX safetensors — Compatible with:** mlx-lm, mlx-vlm, oMLX on...
— Quantization: oQ3 (sensitivity-driven mixed precision, groupsize=64) — Format: MLX safetensors — Compatible with:** mlx-lm, mlx-vlm, oMLX on...