TheBloke/llama-2-13B-chat-limarp-v2-merged-GPTQ - Каталог нейросетей
Генерация текста

TheBloke/llama-2-13B-chat-limarp-v2-merged-GPTQ

Добавлено:
TheBloke/llama-2-13B-chat-limarp-v2-merged-GPTQ

Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported by a grant from andreessen horowitz (a16z) — Model creator: Doctor-Shotgun — Original model: Llama 2 13B Chat — LimaRP v2 Merged This repo contains GPTQ model files for Doctor-Shotgun’s Llama 2 13B Chat — LimaRP v2 Merged. Multiple GPTQ parameter permutations are provided; see Provided Files below for details of the options provided, their parameters, and the software used to create them. AWQ model(s) for GPU inference. GPTQ models for GPU inference, with multiple quantisation parameter options. 2, 3, 4, 5, 6 and 8-bit GGUF models for CPU+GPU inference Doctor-Shotgun’s original unquantised fp16 model in pytorch format, for GPU inference and for further conversions The creator of the source model has listed its license as agpl-3.0, and this quantization has therefore used that same license. As this model is based on Llama 2, it is also subject to the Meta Llama 2 license terms, and the license files for that are additionally included. It should therefore be considered as being claimed to be licensed under both licenses. I contacted Hugging Face…

Модальности:
Генерация текста

Области применения:
Диалог / чат


Задача: Генерация текста
Автор: TheBloke
Теги: llama, llama-2, en, text-generation-inference, 4-bit, gptq
Лайков: 3  |  Загрузок: 28

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.