Qwen3-30B-A3B-Thinking-2507-Eagle3
The NVIDIA Qwen3-30B-A3B-Thinking-2507 Eagle model is the Eagle head of the Alibaba’s Qwen3-30B-A3B-Thinking-2507 model, which is an auto-regressive...
The NVIDIA Qwen3-30B-A3B-Thinking-2507 Eagle model is the Eagle head of the Alibaba’s Qwen3-30B-A3B-Thinking-2507 model, which is an auto-regressive...
NVIDIA-Nemotron-Nano-12B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model...
This is quantized version of nvidia/NVIDIA-Nemotron-Nano-9B-v2 created using llama.cpp NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from...
This model mlx-community/NVIDIA-Nemotron-Nano-9B-v2-4bits was converted to MLX format from nvidia/NVIDIA-Nemotron-Nano-9B-v2 using mlx-lm version 0.26.3. Модальности:Генерация текста Области применения:Диалог...
Llama-3.3-Nemotron-Super-49B-v1.5 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a...
This model was generated using llama.cpp at commit e2b7621e. OpenReasoning-Nemotron-1.5B is a large language model (LLM) which is...
This model was generated using llama.cpp at commit 3f4fc97f. * This is our estimation of the Artificial Analysis...
We introduce the updated version of the Qwen3-235B-A22B non-thinking mode, named Qwen3-235B-A22B-Instruct-2507, featuring the following key enhancements: —...
Llama-3.3-Nemotron-70B-Reward is a large language model that leverages Meta-Llama-3.3-70B-Instruct as the foundation and is fine-tuned using scaled Bradley-Terry...
This is quantized version of nvidia/Llama-3.1-Nemotron-Nano-8B-v1 created using llama.cpp Llama-3.1-Nemotron-Nano-8B-v1 is a large language model (LLM) which is...