Shadows-MoE
Модель генерации текста Модальности:Генерация текста Задача: Генерация текста Автор: Local-Novel-LLM-project Теги: mixtral, mergekit, merge, text-generation-inference, endpoints_compatibleЛайков: 3 | ...
Модель генерации текста Модальности:Генерация текста Задача: Генерация текста Автор: Local-Novel-LLM-project Теги: mixtral, mergekit, merge, text-generation-inference, endpoints_compatibleЛайков: 3 | ...
Dear god, this model took half a day to upload because my internet kept cutting out. Totally worth...
This is a merge of pre-trained language models created using mergekit. The following models were included in the...
This is a merge of pre-trained language models created using mergekit. The following models were included in the...
Original model: https://huggingface.co/sophosympatheia/New-Dawn-Llama-3-70B-32K-v1.0 If the model is bigger than 50GB, it will have been split into multiple files....
This is a merge of pre-trained language models created using mergekit. An expiremental merger inspired by the merger...
L3-8B-Stheno-2x8B-MoE is a Mixture of Experts (MoE) made with the following models using LazyMergekit: Sao10K/L3-8B-Stheno-v3.2 Sao10K/L3-8B-Stheno-v3.1 Модальности:Генерация текста...
New merge method with better results, in all aspects a improvement over the previous version. At the core...
A Llama-3 Decoder only model by combining 2x finetuned Llama-3 70B models into 1. The following models were...
!image/png This is a merge of pre-trained language models created using mergekit. FP32 version — model retains qualities...