Shadows-MoE
Модель генерации текста Модальности:Генерация текста Задача: Генерация текста Автор: Local-Novel-LLM-project Теги: mixtral, mergekit, merge, text-generation-inference, endpoints_compatibleЛайков: 3 | ...
Модель генерации текста Модальности:Генерация текста Задача: Генерация текста Автор: Local-Novel-LLM-project Теги: mixtral, mergekit, merge, text-generation-inference, endpoints_compatibleЛайков: 3 | ...
— GGUF version of KARAKURI LM 8x7B Instruct v0.1 — Debeloped by: KARAKURI Inc. — Languages: Primarily English...
L3-8B-Stheno-2x8B-MoE is a Mixture of Experts (MoE) made with the following models using LazyMergekit: Sao10K/L3-8B-Stheno-v3.2 Sao10K/L3-8B-Stheno-v3.1 Модальности:Генерация текста...
— Model Architecture: Mixtral-8x22B-Instruct-v0.1 — Input: Text — Output: Text — Model Optimizations: — Weight quantization: FP8 —...
Exllamav2 3.75bpw quantization of Typhon-Mixtral-v1 from Sao10K, quantized with default calibration dataset. > [!IMPORTANT] >This bpw is the...
The Mistral-DNA-v1-422M-hg38 Large Language Model (LLM) is a pretrained generative DNA sequence model with 422M parameters. It is...
> [!WARNING] > This model checkpoint is provided as-is and might not be up-to-date. Please use the corresponding...
The Mistral-Chem-v1-417M Large Language Model (LLM) is a pretrained generative chemical molecule model with 417M parameters. It is...
This is a merge of pre-trained language models created using mergekit. This model is a merge of all...
I’ve been really impressed with how well these frankenmoe models quant compared to the base llama 8b, but...