Capella-Qwen3-DS-V3.1-4B
> Capella-Qwen3-DS-V3.1-4B is a reasoning-focused model fine-tuned on Qwen-4B using DeepSeek v3.1 synthetic traces (10K). > It specializes...
> Capella-Qwen3-DS-V3.1-4B is a reasoning-focused model fine-tuned on Qwen-4B using DeepSeek v3.1 synthetic traces (10K). > It specializes...
This repo contains the full precision source code, in «safe tensors» format to generate GGUFs, GPTQ, EXL2, AWQ,...
This repository hosts the Phi4-mini-instruct model quantized with torchao using int4 weight-only quantization and the awq algorithm. This...
This model Qwen3-42B-A3B-2507-Thinking-Abliterated-uncensored-TOTAL-RECALL-v2-Medium-MASTER-CODER-qx4-mlx was converted to MLX format from DavidAU/Qwen3-42B-A3B-2507-Thinking-Abliterated-uncensored-TOTAL-RECALL-v2-Medium-MASTER-CODER using mlx-lm version 0.26.3. Модальности:Генерация текста Области применения:Генерация...
Devstral Small 2507 is a powerful choice for local inference, achieving SOTA open source results at just 24B...
Jimi is a LLM fine-tuned variant of Google Deepmind’s gemma-4-E4B-it transformer model, optimized for enhanced contextual comprehension, instruction...
This model was generated using llama.cpp at commit e2b7621e. OpenReasoning-Nemotron-1.5B is a large language model (LLM) which is...
This model was generated using llama.cpp at commit 3f4fc97f. * This is our estimation of the Artificial Analysis...
Original model: https://huggingface.co/microsoft/NextCoder-32B Run them directly with llama.cpp, or any other llama.cpp based project Some of these quants...
AWQ quantization: done by stelterlab in INT4 GEMM with AutoAWQ by casper-hansen (https://github.com/casper-hansen/AutoAWQ/) Original Weights by Qwen AI/Finetuned...