— Homepage: https://lklab.kaist.ac.kr/Janus/ — Repository: https://github.com/kaistAI/Janus — Paper: https://arxiv.org/abs/2405.17977 — Point of Contact: seongyun@kaist.ac.kr Janus is a model trained using Mistral-7B-v0.2 as its base model. Janus has been trained on Multifaceted Collection, a preference dataset containing 196k unique system messages for aligning LLMs to diverse human preferences. Janus not only excels at generating personalized responses that cater to various human preferences but is also adept at producing responses that are generally preferred for being helpful and harmless. Janus-DPO-7B is a model created by applying DPO to Janus using the Multifaceted-Collection-DPO. — Model type: Language model — Language(s) (NLP): English — License: Apache 2.0 — Related Models: Janus-7B, Janus-ORPO-7B, Janus-RM-7B — Training Datasets: Multifaceted-Collection-DPO — Resources for more information: — Research paper — GitHub Repo Janus is a model generalized for various system messages, allowing users to control the model’s response by inputting the desired system message. The input prompt format is as follows: Additionally, an example of the inference code applying this is as…
Модальности:
Генерация текста
Задача: Генерация текста
Автор: kaist-ai
Теги: tensorboard, mistral, axolotl, trl, generated_from_trainer, dpo, en, text-generation-inference
Лайков: 3 | Загрузок: 8,574
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.