STEVE-13b
Zhonghan Zhao1 , Wenhao Chai2❤, Xuan Wang1*, Li Boyi1, Shengyu Hao1, Shidong Cao1, Tian Ye3, Jenq-Neng Hwang2, Gaoang...
Zhonghan Zhao1 , Wenhao Chai2❤, Xuan Wang1*, Li Boyi1, Shengyu Hao1, Shidong Cao1, Tian Ye3, Jenq-Neng Hwang2, Gaoang...
Model Architecture PlatYi-34B-Llama-Q-v3 is an auto-regressive language model based on the Yi-34B transformer architecture. — Before model, there...
This is an OpenLlama model that has been fine-tuned on 1 epoch of the AlpacaCode dataset (122K rows)....
Dolphin-2.2-yi-34b-200k, Nous-Capybara-34B, Tess-M-v1.4, Airoboros-31-yi-34b-200k, PlatYi-34B-200K-Q, and Una-xaberius-34b-v1beta** merged with a new, experimental implementation of «dare ties» via mergekit....
Idea is an updated version of Euryale with ReMantik instead of the ties-merge between the original 3 models....
This is a Llama 2-based model consisting of a merge between: — Sao10K/Euryale-1.3-L2-70B — allenai/tulu-2-dpo-70b — GOAT-AI/GOAT-70B-Storytelling —...
We create various AI models and develop solutions that can be applied to businesses. And as for generative...
This is an ShearedPlats-7b model that has been fine-tuned on 2 epochs of the Open-Platypus dataset. Модальности:Генерация текста...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
EAGLE (Extrapolation Algorithm for Greater Language-model Efficiency) is a new baseline for fast decoding of Large Language Models...