joshmiller656/Llama-3.1-Nemotron-70B-Instruct-AWQ-INT4 - Каталог нейросетей
Генерация текста

joshmiller656/Llama-3.1-Nemotron-70B-Instruct-AWQ-INT4

Добавлено:
joshmiller656/Llama-3.1-Nemotron-70B-Instruct-AWQ-INT4

If you run into errors on a multi GPU machine, I’ve found that setting CUDAVISIBLEDEVICES=0 helps. Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries. INVALID LANGUAGE PAIR SPECIFIED. EXAMPLE: LANGPAIR=EN|IT USING 2 LETTER ISO OR RFC3066 LIKE ZH-CN. ALMOST ALL LANGUAGES SUPPORTED BUT SOME MAY HAVE NO CONTENT This model was trained using RLHF (specifically, REINFORCE), Llama-3.1-Nemotron-70B-Reward and HelpSteer2-Preference prompts on a Llama-3.1-70B-Instruct model as the initial policy. Llama-3.1-Nemotron-70B-Instruct-HF has been converted from Llama-3.1-Nemotron-70B-Instruct to support it in the HuggingFace Transformers codebase. Please note that evaluation results might be slightly different from the Llama-3.1-Nemotron-70B-Instruct as evaluated in NeMo-Aligner, which the…

Модальности:
Генерация текста

Области применения:
Диалог / чат Следование инструкциям


Задача: Генерация текста
Автор: joshmiller656
Теги: llama, nemotron, awq, quantized, int4, conversational, en, de
Лайков: 3  |  Загрузок: 110

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.