Mixtral-8x7B-Instruct-v0.1-limarp
Experimental model, using a limarp qlora trained at 10k ctx length (greater than size of the longest limarp...
Experimental model, using a limarp qlora trained at 10k ctx length (greater than size of the longest limarp...
This repository contains Chinese-LLaMA-2-7B-64K, which is tuned on Chinese-LLaMA-2-7B with YaRN method. For LoRA-only model, please see: https://huggingface.co/hfl/chinese-llama-2-lora-7b-64k...
Our finetuned Mistral LLM is a large language model specialized for natural language processing tasks, delivering enhanced performance...
augmxnt/shisa-7b-v1 Mistral-7B base Pre-trained on 8B of MADLAD-Ja Finetuned on Japanese instructions Highest scoring 7B model on conversation...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
LinguaMatic is an advanced AI model designed to handle a wide range of Natural Language Processing (NLP) tasks....
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
— Introduction — Diff of TransNormerLLM2 — Released Weights — Benchmark Results — Inference and Deployment — Dependency...
— Introduction — Diff of TransNormerLLM2 — Released Weights — Benchmark Results — Inference and Deployment — Dependency...
The model is the GGUF version of rinna/nekomata-14b. It can be used with llama.cpp for lightweight inference. Quantization...