comma-v0.1-2t-GGUF
This is a GGUF conversion of common-pile/comma-v0.1-2t for use with llama.cpp and Ollama. Original Model: Comma v0.1-2T Architecture:...
This is a GGUF conversion of common-pile/comma-v0.1-2t for use with llama.cpp and Ollama. Original Model: Comma v0.1-2T Architecture:...
static quants of https://huggingface.co/mookiezii/Discord-Hermes-3-8B For a convenient overview and download list, visit our model page for this model....
weighted/imatrix quants of https://huggingface.co/mookiezi/Discord-Micae-Hermes-3-3B For a convenient overview and download list, visit our model page for this model....
Llama-3.3-Nemotron-Super-49B-v1.5 is a significantly upgraded version of Llama-3.3-Nemotron-Super-49B-v1 and is a large language model (LLM) which is a...
KunoRZN-Llama-3-3B (Knowledge Understanding Network with Optimized Reasoning Zone Navigation) is VinkuraAI’s flagship language model designed to support 12+...
Наш последний метод квантования вводит прецизионно-адаптивное квантование для сверхнизкоразрядных моделей (1-2 бита) с проверенными улучшениями на Llama-3-8B. Этот...
This is quantized version of nvidia/Llama-3.1-Nemotron-Nano-8B-v1 created using llama.cpp Llama-3.1-Nemotron-Nano-8B-v1 is a large language model (LLM) which is...
This model was generated using llama.cpp at commit f5cd27b7. Our latest quantization method introduces precision-adaptive quantization for ultra-low-bit...
Наш последний метод квантования вводит прецизионно-адаптивное квантование для сверхнизкоразрядных моделей (1-2 бита) с проверенными улучшениями на Llama-3-8B. Этот...
We have a free Google Colab notebook for turning Llama 3.1 (8B) into a reasoning model: https://colab.research.google.com/github/unslothai/notebooks/blob/main/nb/Llama3.1_(8B)-GRPO.ipynb All...