allenai/OLMo-2-1124-7B-RM - Каталог нейросетей
Генерация текста

allenai/OLMo-2-1124-7B-RM

Добавлено:
allenai/OLMo-2-1124-7B-RM

Upon the initial release of OLMo-2 models, we realized the post-trained models did not share the pre-tokenization logic that the base models use. As a result, we have trained new post-trained models. The new models are available under the same names as the original models, but we have made the old models available with a postfix «-preview». See OLMo 2 Preview Post-trained Models for the colleciton of the legacy models. OLMo-2 7B RM November 2024 is post-trained variant of the OLMo-2 13B November 2024 model, which has undergone supervised finetuning on an OLMo-specific variant of the Tülu 3 dataset and further DPO training on this dataset. Tülu 3 is designed for state-of-the-art performance on a diversity of tasks in addition to chat, such as MATH, GSM8K, and IFEval. Check out the OLMo 2 paper or Tülu 3 paper for more details! OLMo is a series of Open Language Models designed to enable the science of language models. These models are trained on the Dolma dataset. We are releasing all code, checkpoints, logs (coming soon), and associated training details. The core models released in this batch include the following: — Model type: A model trained on a mix of publicly available,…

Модальности:
Генерация текста

Области применения:
Диалог / чат


Задача: Генерация текста
Автор: allenai
Теги: olmo2, text-classification, conversational, en, endpoints_compatible
Лайков: 3  |  Загрузок: 1,255

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.