Athene-70B-IMat-GGUF
> [!IMPORTANT] > This repo is quantized from the updated version of Athene-70B as of the 24th of...
> [!IMPORTANT] > This repo is quantized from the updated version of Athene-70B as of the 24th of...
Original Model: facebook/llm-compiler-13b Original dtype: BF16 (bfloat16) Quantized by: llama.cpp b3256 IMatrix dataset: here — Files — IMatrix...
Llama.cpp imatrix quantization of OpenLLM-Ro/RoLlama3-8b-Instruct Original Model: OpenLLM-Ro/RoLlama3-8b-Instruct Original dtype: BF16 (bfloat16) Quantized by: llama.cpp b3206 IMatrix dataset:...
Оригинальная модель: Qwen/Qwen2-7B-Instruct Оригинальный dtype: BF16 (bfloat16) Quantized by: llama.cpp b3091 IMatrix dataset: here — Files — IMatrix...
Original Model: NTQAI/Nxcode-CQ-7B-orpo Original dtype: BF16 (bfloat16) Quantized by: llama.cpp b3067 IMatrix dataset: here — Files — IMatrix...
Original Model: LLM360/K2 Original dtype: FP16 (float16) Quantized by: llama.cpp b3051 IMatrix dataset: here — Files — IMatrix...
Original Model: IEITYuan/Yuan2-M32-hf Original dtype: BF16 (bfloat16) Quantized by: https://github.com/chong000/3rd_party/tree/main IMatrix dataset: here — Files — IMatrix —...
Llama.cpp imatrix quantization of internlm/internlm2-math-plus-18b_ Original Model: internlm/internlm2-math-plus-18b Original dtype: BF16 (bfloat16`) Quantized by: llama.cpp b3008 IMatrix dataset:...
Llama 2 (13B) fine-tuned on Clibrain’s Spanish instructions dataset and optimized using GPTQ. Llama 2 is a collection...
Версия GPTQ модели Open-Assistant StableLM-7B SFT-7, квантованная с помощью AutoGPTQ Обновит это позже — взгляните на солаб, чтобы...