Llama3.1-70B-ShiningValiant2-GGUF
Original model: https://huggingface.co/ValiantLabs/Llama3.1-70B-ShiningValiant2 Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
Original model: https://huggingface.co/ValiantLabs/Llama3.1-70B-ShiningValiant2 Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
Это квантованная версия ValiantLabs/Llama3.2-3B-Esper2, созданная с использованием llama.cpp Esper 2 — специалист по DevOps и коду облачной архитектуры,...
Это квантованная версия ValiantLabs/Llama3.2-3B-ShiningValiant2, созданная с использованием llama.cpp Shining Valiant 2 — это модель чата, построенная на Llama...
Esper 2 is a DevOps and cloud architecture code specialist built on Llama 3.2 3b. — Expertise-driven, an...
Original model: https://huggingface.co/ValiantLabs/Llama3.1-8B-Enigma Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
Fireplace is a function-calling model for Llama 3 70b Instruct. — combines function-calling abilities with a high-performance, versatile...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Guardpoint: gemma-4-12B, Qwen3-14B, gpt-oss-20b, Qwen3.5-27B, Qwen3-32B, gpt-oss-120b Guardpoint is a medical reasoning specialist built on gpt-oss. — Finetuned...
Guardpoint: gemma-4-12B, Qwen3-14B, gpt-oss-20b, Qwen3.5-27B, Qwen3-32B, gpt-oss-120b Guardpoint is a medical reasoning specialist built on Qwen 3. —...
Esper 3.1: Ministral-3-3B-Reasoning-2512, Qwen3-4B-Thinking-2507, Ministral-3-8B-Reasoning-2512, Ministral-3-14B-Reasoning-2512, gpt-oss-20b, Qwen3.5-27B, Qwen3.6-27B, Qwen3.6-35B-A3B Esper 3.1 is a coding, architecture, and DevOps...