unsloth/Llama-3.1-Storm-8B - Каталог нейросетей
Генерация текста

unsloth/Llama-3.1-Storm-8B

Добавлено:
unsloth/Llama-3.1-Storm-8B

We have a free Google Colab Tesla T4 notebook for Llama 3.1 (8B) here: https://colab.research.google.com/drive/1Ys44kVvmeZtnICzWz0xgpRnrIOjZAuxp?usp=sharing All notebooks are beginner friendly! Add your dataset, click «Run All», and you’ll get a 2x faster finetuned model which can be exported to GGUF, vLLM or uploaded to Hugging Face. Authors: Ashvini Kumar Jindal, Pawan Kumar Rajpoot, Ankur Parikh, Akshita Sukhlecha 🤗 Hugging Face Announcement Blog: https://huggingface.co/blog/akjindal53244/llama31-storm8b We present the Llama-3.1-Storm-8B model that outperforms Meta AI’s Llama-3.1-8B-Instruct and Hermes-3-Llama-3.1-8B models significantly across diverse benchmarks as shown in the performance comparison plot in the next section. Our approach consists of three key steps: 1. Self-Curation: We applied two self-curation methods to select approximately 1 million high-quality examples from a pool of ~2.8 million open-source examples. Our curation criteria focused on educational value and difficulty level, using the same SLM for annotation instead of larger models (e.g. 70B, 405B). 2. Targeted fine-tuning: We performed Spectrum-based targeted fine-tuning over the Llama-3.1-8B-Instruct…

Модальности:
Генерация текста

Области применения:
Диалог / чат


Задача: Генерация текста
Автор: unsloth
Теги: llama, llama-3, meta, facebook, unsloth, conversational, en, text-generation-inference
Лайков: 3  |  Загрузок: 70

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.