gemma-4-31B-it-qat-NVFP4-Blackwell
Этот репозиторий содержит настраиваемую по инструкциям модель Gemma 4 31B, квантованную с нативной точностью FP4 (NVFP4) для высокоэффективного...
Этот репозиторий содержит настраиваемую по инструкциям модель Gemma 4 31B, квантованную с нативной точностью FP4 (NVFP4) для высокоэффективного...
This model was converted to MLX format from nvidia/Nemotron-Labs-Diffusion-3B using mlx-vlm version 0.6.3. Refer to the original model...
Nemotron-Cascade-2-30B-A3B ships with custom model code (modelingnemotronh.py) that has two bugs exposed by causalconv1d 1.6.x. These are bugs...
Abliterated version of openNemo-9B with safety refusals removed. Built using Snakehead — Empero AI’s internal abliteration tool specialized...
AWQ quantization: done by stelterlab in INT4 with llm-compressor (https://github.com/vllm-project/llm-compressor — v0.10.0.1) from the vllm-project. See recipe.yaml for...
We’re excited to introduce Nemotron-Cascade-2-30B-A3B, an open 30B MoE model with 3B activated parameters that delivers strong reasoning...
All evaluations were done using NeMo-Skills. We published a tutorial with all details necessary to reproduce our evaluation...
The post-training data has a cutoff date of November 28, 2025. The pre-training data has a cutoff date...
This model mlx-community/NVIDIA-Nemotron-3-Nano-30B-A3B-4bit was converted to MLX format from nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16 using mlx-lm version 0.29.0. Модальности:Генерация текста Области применения:Диалог...
> [!NOTE] > Includes Unsloth chat template fixes! For llama.cpp, use —jinja > Unsloth Dynamic 2.0 achieves superior...