Enjoy! Raise an issue if you’d like other BPW levels. Base Model Card Follows: —————————|:—————-|:——|:———-|:——-| You can use the model using HuggingFace Transformers library with 2 or more 80GB GPUs (NVIDIA Ampere or newer) with at least 150GB of free disk space to accomodate the download. This code has been tested on Transformers v4.44.0, torch v2.4.0 and 2 A100 80GB GPUs, but any setup that supports should support this model as well. If you run into problems, you can consider doing . If you find this model useful, please cite the following works HelpSteer2-Preference SteerLM method HelpSteer HelpSteer2 Introducing Llama 3.1: Our most capable models to date Meta’s Llama 3.1 Webpage * Meta’s Llama 3.1 Model Card Architecture Type: Transformer Network Architecture: Llama 3.1 Input Type(s): Text Input Format: String Input Parameters: One Dimensional (1D) Other Properties Related to Input: Max of 128k tokens Output Type(s): Text Output Format: String
Модальности:
Генерация текста
Области применения:
Диалог / чат Следование инструкциям
Задача: Генерация текста
Автор: bigstorm
Теги: llama, nvidia, llama3.1, conversational, en, text-generation-inference, exl2
Лайков: 3 | Загрузок: 25
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.