Ahma-7B-Instruct is a instruct/chat-tuned version of Ahma-7B trained to follow instructions in Finnish. The base Ahma 7B parameter model is decoder-only transformer model based on Meta’s Llama (v1) architecture pretrained from scratch on Finnish language. Original Llama model architecture was introduced in this paper and first released at this page. What does Ahma mean? Ahma is the Finnish word for wolverine! In the Finnish Lapland, wolverines are the biggest cause of reindeer damage. There are two different sized base Ahma models, all pretrained from scratch for 139B tokens: This model was fine-tuned for instruction following. Instruction-tuned models are intended for assistant-like chat, whereas pretrained models can be adapted for a variety of natural language generation tasks. If you want to use this model for instruction-following, you need to use the same prompt format we used in the fine-tuning process (basically the same format what Meta used in their Llama2 models). Note: do not use «LlamaTokenizer» from transformers library but always use the AutoTokenizer instead, or use the plain sentencepiece tokenizer. Here is an example using the instruction-following prompt format…
Модальности:
Генерация текста
Области применения:
Диалог / чат Следование инструкциям
Задача: Генерация текста
Автор: Finnish-NLP
Теги: llama, finnish, conversational, fi
Лайков: 3 | Загрузок: 80
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.