This model is a merged pre-trained language model created using MergeKit with the TIES merge method. It uses Qwen/Qwen2.5-7B-Instruct-1M as the base and combines deepseek-ai/DeepSeek-R1-Distill-Qwen-7B and Qwen/Qwen2.5-7B-Instruct with equal weight and density. The merge configuration includes normalization, int8 masking, and bfloat16 precision for optimized performance. This is a merge of pre-trained language models created using mergekit. This model was merged using the TIES merge method using Qwen/Qwen2.5-7B-Instruct-1M as a base. The following models were included in the merge: deepseek-ai/DeepSeek-R1-Distill-Qwen-7B Qwen/Qwen2.5-7B-Instruct The following YAML configuration was used to produce this model:
Модальности:
Генерация текста
Области применения:
Диалог / чат Логика и рассуждение
Задача: Генерация текста
Автор: prithivMLmods
Теги: qwen2, mergekit, merge, conversational, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 58
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.