Llama3-8B-1.58-100B-tokens-GGUF
The Llama3-8B-1.58 models are large language models fine-tuned on the BitNet 1.58b architecture, starting from the base model...
The Llama3-8B-1.58 models are large language models fine-tuned on the BitNet 1.58b architecture, starting from the base model...
— Model creator: mistralai — Original model: mistralai/Mistral-Small-Instruct-2409 MaziyarPanahi/Mistral-Small-Instruct-2409-GGUF contains GGUF format model files for mistralai/Mistral-Small-Instruct-2409. GGUF is...
Enjoy! Raise an issue if you’d like other BPW levels. Base Model Card Follows: —————————|:—————-|:——|:———-|:——-| You can use...
Qwen2.5 — это последняя серия больших языковых моделей Qwen. Для Qwen2.5 мы выпускаем ряд базовых языковых моделей и...
— Model creator: deepseek-ai — Original model: deepseek-ai/DeepSeek-V2.5 MaziyarPanahi/DeepSeek-V2.5-GGUF contains GGUF format model files for deepseek-ai/DeepSeek-V2.5. GGUF is...
This model was obtained by quantizing the weights and activations of Bielik-11B-v.2.2-Instruct to W8A8 (INT8) data type, ready...
This model was converted to Quanto format from SpeakLeash’s Bielik-11B-v.2.2-Instruct. DISCLAIMER: Be aware that quantised models show reduced...
This is 8-bit GPTQ version of Meta-Llama-3.1-8B-Instruct. Quantization has been done using AutoGPTQ library. Starting with transformers >=...
— Архитектура модели: Meta-Llama-3 — Ввод: Текст — Вывод: Текст — Оптимизация модели: — Квантование веса: INT8 —...
Trained with compute from Backyard.ai | Thanks to them and @dynafire for helping me out. Relevant Axolotl Configurations:...