astronomer/Llama-3-8B-GPTQ-8-Bit - Каталог нейросетей
Генерация текста

astronomer/Llama-3-8B-GPTQ-8-Bit

Добавлено:
astronomer/Llama-3-8B-GPTQ-8-Bit

This model is generously created and made open source by Astronomer. Astronomer is the de facto company for Apache Airflow, the most trusted open-source framework for data orchestration and MLOps. — Original Model creator: Meta Llama from Meta — Original model: meta-llama/Meta-Llama-3-8B — Built with Meta Llama 3 — Quantized by David Xue from Astronomer — If you intend to fine-tune this model with any added tokens, or fine-tune for instruction following, please use the untrained-special-tokens-fixed branch/revision. — Special tokens such as the ones used for instruct are undertrained in Llama 3 base models. — Credits: discovered by Daniel Han https://twitter.com/danielhanchen/status/1781395882925343058 — !image/png — For loading this model onto vLLM, make sure all requests have «stoptokenids»:[128001, 128009] to temporarily address the non-stop generation issue. — vLLM does not yet respect generationconfig.json. — vLLM team is working on a a fix for this https://github.com/vllm-project/vllm/issues/4180 — For oobabooga/text-generation-webui — Load the model via AutoGPTQ, with noinjectfusedattention enabled. This is a bug with AutoGPTQ library. — Under Parameters -> Generation ->…

Модальности:
Генерация текста


Задача: Генерация текста
Автор: astronomer
Теги: llama, llama-3, facebook, meta, astronomer, gptq, pretrained, quantized
Лайков: 3  |  Загрузок: 22

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.