cybertron-v4-qw7B-MGS-GGUF
Original model: https://huggingface.co/fblgit/cybertron-v4-qw7B-MGS Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
Original model: https://huggingface.co/fblgit/cybertron-v4-qw7B-MGS Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
— Developed by: TroyDoesAI — License: apache-2.0 — Finetuned from model : TroyDoesAI/BlackSheep-Llama3.2-3B This llama model was trained...
Optimized GGUF quantization files for enhanced model performance Модальности:Генерация текста Области применения:Диалог / чат Задача: Генерация текста Автор:...
If you run into errors on a multi GPU machine, I’ve found that setting CUDAVISIBLEDEVICES=0 helps. Llama-3.1-Nemotron-70B-Instruct is...
Optimized GGUF quantization files for enhanced model performance Модальности:Генерация текста Области применения:Диалог / чат Задача: Генерация текста Автор:...
Original model: https://huggingface.co/Bllossom/llama-3.2-Korean-Bllossom-3B Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
della_linear merge done at a 40/60 split using Qwen2.5-32B-Instruct and EVA-Qwen2.5-32B-v0.0. Seems pretty good at creative ventures so...
Original model: https://huggingface.co/TheDrummer/Nautilus-70B-v0.1 Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
The Meta Llama 3.2 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction-tuned...
This is a merge of pre-trained language models created using mergekit. This model was merged using the TIES...