Qwythos-9B-Claude-Mythos-5-1M-NVFP4
NVFP4 (NVIDIA FP4, weight-only NVFP4A16) quantization of empero-ai/Qwythos-9B-Claude-Mythos-5-1M — a Claude-Mythos/Fable-trace reasoning fine-tune of Qwen3.5-9B (qwen35: a dense...
NVFP4 (NVIDIA FP4, weight-only NVFP4A16) quantization of empero-ai/Qwythos-9B-Claude-Mythos-5-1M — a Claude-Mythos/Fable-trace reasoning fine-tune of Qwen3.5-9B (qwen35: a dense...
NVFP4-quantized build of prefeitura-rio/Rio-3.5-Open-397B — a 397B-parameter (17B active) Qwen3.5-MoE vision-language model (512 experts, hybrid softmax + linear/DeltaNet...
Этот репозиторий содержит черновик модели Gemma 4 31B Instruction-Tuned Assistant, квантованный до собственной точности FP4 (NVFP4) для высокоэффективного...
Этот репозиторий содержит настраиваемую по инструкциям модель Gemma 4 31B, квантованную с нативной точностью FP4 (NVFP4) для высокоэффективного...
NVFP4 (W4A4) quantization of huihui-ai/Huihui-LFM2.5-8B-A1B-abliterated — the abliterated (refusal-reduced) build of Liquid AI’s 8.3B-total / 1.5B-active mixture-of-experts reasoner...
What happened: The initial upload (2026-04-15) used ignore=[«lmhead»] in the llm-compressor recipe, which meant the 62 MoE routers...
— Model Architecture: Qwen/Qwen3-235B-A22B-Instruct-2507 — Input: Text — Output: Text — Model Optimizations: — Weight quantization: FP4 —...
We introduce the updated version of the Qwen3-235B-A22B non-thinking mode, named Qwen3-235B-A22B-Instruct-2507, featuring the following key enhancements: —...
The test results in the following table are based on the MMLU benchmark. In order to speed up...
Модель генерации текста Модальности:Генерация текста Области применения:Диалог / чат Задача: Генерация текста Автор: RazielXYZ Теги: qwen3_5_moe, quantized, fp4,...