flammenai/Mahou-1.1-llama3-8B finetuned on a Japanese DPO set. 1. Use ChatML for the Context Template. 2. Turn on Instruct Mode for ChatML. 3. Use the following stopping strings: [» This model is based on Meta Llama-3-8B and is governed by the META LLAMA 3 COMMUNITY LICENSE AGREEMENT. Fine-tune a Mistral-7b model with Direct Preference Optimization — Maxime Labonne
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: nbeerbower
Теги: llama, conversational, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 22
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.