Please note that access is limited to students, companies, and organizations from Nordic countries. Kindly provide your work email or student email to access the models. Thank you for your understanding. NowAI LLM is intended for both commercial and research use in Nordic countries. To get access to the model, please carefully read the message and complete the required information. The model may have potential risks common to large language models, such as hallucination, factual inconsistency, toxicity, and bias etc. All NorwAI LLM models were (continue-)pretrained on 51.15 Billion tokens, equivalent to 30.33 Billion words, sourced from public datasets and data shared by Schibsted, NRK, and VG partners under agreement. The publicly available datasets were preprocessed to filter out texts with copyright issues, and all datasets were preprocessed to remove sensitive information. Among all the pretraining data, the ratio of Norwegian to other languages is 3 to 2. Among the languages other than Norwegian, Swedish and Danish account for the majority, with a small amount of English and other languages. All models were pretrained and fine-tuned using the llm-foundary framework on the…
Модальности:
Генерация текста
Задача: Генерация текста
Автор: NorwAI
Теги: mistral, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 0
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.