OmniLing-V1-8b-experimental-GGUF
This is quantized version of WoonaAI/OmniLing-V1-8b-experimental created using llama.cpp OmniLing — модель, созданная для переводов между русским и...
This is quantized version of WoonaAI/OmniLing-V1-8b-experimental created using llama.cpp OmniLing — модель, созданная для переводов между русским и...
— Model creator: jinaai — Original model: jinaai/reader-lm-1.5b arcee-train/reader-lm-1.5b-GGUF contains GGUF format model files for jinaai/reader-lm-1.5b. GGUF is...
Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings and output weights...
Original model: https://huggingface.co/deepseek-ai/DeepSeek-Coder-V2-Instruct-0724 Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings...
— Model creator: deepseek-ai — Original model: deepseek-ai/DeepSeek-V2.5 MaziyarPanahi/DeepSeek-V2.5-GGUF contains GGUF format model files for deepseek-ai/DeepSeek-V2.5. GGUF is...
 This is quantized version of ibm-granite/granite-8b-code-instruct-128k created using llama.cpp Granite-8B-Code-Instruct-128K is a 8B parameter long-context instruct model...
 This is quantized version of lelapa/InkubaLM-0.4B created using llama.cpp InkubaLM has been trained from scratch using 1.9...
 This is quantized version of bigcode/starcoder2-3b created using llama.cpp 1. Model Summary 2. Use 3. Limitations 4....
This is an experimental version of the repository containing quantized Bielik-11B-v.2.3-Instruct models using calibration with importance matrix (imatrix)....
Some of these quants (Q3KXL, Q4KL etc) are the standard quantization method with the embeddings and output weights...