We have a Qwen 2.5 (all model sizes) free Google Colab Tesla T4 notebook. Also a Qwen 2.5 conversational style notebook. All notebooks are beginner friendly! Add your dataset, click «Run All», and you’ll get a 2x faster finetuned model which can be exported to GGUF, vLLM or uploaded to Hugging Face. — This conversational notebook is useful for ShareGPT ChatML / Vicuna templates. — This text completion notebook is for raw text. This DPO notebook replicates Zephyr. — * Kaggle has 2x T4s, but we use 1. Due to overhead, 1x T4 is 5x faster. Qwen2.5 is the latest series of Qwen large language models. For Qwen2.5, we release a number of base language models and instruction-tuned language models ranging from 0.5 to 72 billion parameters. Qwen2.5 brings the following improvements upon Qwen2: — Significantly more knowledge and has greatly improved capabilities in coding and mathematics, thanks to our specialized expert models in these domains. — Significant improvements in instruction following, generating long texts (over 8K tokens), understanding structured data (e.g, tables), and generating structured outputs especially JSON. More resilient to the diversity of system prompts, enhancing…
Модальности:
Генерация текста
Задача: Генерация текста
Автор: unsloth
Теги: qwen2, unsloth, zho, eng, fra, spa, por, deu
Лайков: 3 | Загрузок: 539
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.