Document Title h1 { font-size: 36px; color: navy; font-family: ‘Tahoma’; text-align: center; } Introducing the Kancil family of open models Kancil is a fine-tuned version of Llama 3 8B using synthetic QA dataset generated with Llama 3 70B. Version zero of Kancil is the first generative Indonesian LLM gain functional instruction performance using solely synthetic data. ❕Go straight to the colab demo❕ Beta preview I am ultra-overjoyed to introduce you… the 🦌 Kancil! It’s a fine-tuned version of Llama 3 8B with the Tumpeng, an instruction dataset of 14.8 million words. Both the model and dataset is openly available in Huggingface. 📚 The dataset was synthetically generated from Llama 3 70B. A big problem with existing Indonesian instruction dataset is they’re in reality not-very-good-translations of English datasets. Llama 3 70B can generate fluent Indonesian! (with minor caveats 😔) 🦚 This follows previous efforts for collection of open, fine-tuned Indonesian models, like Merak and Cendol. However, Kancil solely leverages synthetic data in a very creative way, which makes it a very unique contribution! This is the second working prototype, Kancil V1. ✨ Training — 2.2x Dataset…
Модальности:
Генерация текста
Задача: Генерация текста
Автор: afrizalha
Теги: llama, unsloth, llama3, indonesia, id, text-generation-inference
Лайков: 3 | Загрузок: 32
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.