nytheria-3b
Nytheria-3B is a masterfully fine-tuned version of the Qwen/Qwen2.5-3B-Instruct model, engineered from the ground up to be a...
Nytheria-3B is a masterfully fine-tuned version of the Qwen/Qwen2.5-3B-Instruct model, engineered from the ground up to be a...
This model has been specially optimized to improve the performance of quantized inference and is recommended for use...
Qwen/Qwen2.5-0.5B-Instruct, but with the vocab of microsoft/phi-4 transplanted using transplant-vocab. Made from the instruct qwen to be used...
Original model: https://huggingface.co/alamios/Mistral-Small-3.1-DRAFT-0.5B Run them directly with llama.cpp, or any other llama.cpp based project Some of these quants...
This model is meant to be used as draft model for speculative decoding with mistralai/Mistral-Small-3.1-24B-Instruct-2503 or mistralai/Mistral-Small-24B-Instruct-2501 The...
This model is trained on outputs of deepseek-ai/DeepSeek-R1-Distill-Qwen-32B and is meant to be used only as draft model...
COCO-7B-Instruct [ chain of continuesness ] is based on a 7B-parameter architecture, optimized for instruction-following tasks and advanced...
QWQ R1 [Reasoning] Distill 1.5B CoT is a fine-tuned language model designed for advanced reasoning and instruction-following tasks....
This is a merge of pre-trained language models created using mergekit. This model was merged using the TIES...
The Novaeus-Promptist-7B-Instruct is a fine-tuned large language model derived from the Qwen2.5-7B-Instruct base model. It is optimized for...