We have a free Google Colab Tesla T4 notebook for Llama 3.1 (8B) here: https://colab.research.google.com/github/unslothai/notebooks/blob/main/nb/Llama3.1_(8B)-Alpaca.ipynb All notebooks are beginner friendly! Add your dataset, click «Run All», and you’ll get a 2x faster finetuned model which can be exported to GGUF, vLLM or uploaded to Hugging Face. — This Llama 3.2 conversational notebook-Conversational.ipynb) is useful for ShareGPT ChatML / Vicuna templates. — This text completion notebook-TextCompletion.ipynb) is for raw text. This DPO notebook replicates Zephyr. — Kaggle has 2x T4s, but we use 1. Due to overhead, 1x T4 is 5x faster. A huge thank you to the DeepSeek team for creating and releasing these models. We introduce our first-generation reasoning models, DeepSeek-R1-Zero and DeepSeek-R1. DeepSeek-R1-Zero, a model trained via large-scale reinforcement learning (RL) without supervised fine-tuning (SFT) as a preliminary step, demonstrated remarkable performance on reasoning. With RL, DeepSeek-R1-Zero naturally emerged with numerous powerful and interesting reasoning behaviors. However, DeepSeek-R1-Zero encounters challenges such as endless repetition, poor readability,…
Модальности:
Генерация текста
Области применения:
Диалог / чат Логика и рассуждение
Задача: Генерация текста
Автор: unsloth
Теги: qwen2, deepseek, qwen, unsloth, conversational, en, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 2,087
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.