adamo1139/Yi-6B-200K-AEZAKMI-v2-6bpw-exl2 - Каталог нейросетей
Генерация текста

adamo1139/Yi-6B-200K-AEZAKMI-v2-6bpw-exl2

Добавлено:
adamo1139/Yi-6B-200K-AEZAKMI-v2-6bpw-exl2

Yi-6B 200K base model fine-tuned on AEZAKMI v2 dataset. It’s like airoboros but hopefully with less gptslop, no refusals and less typical language used by RLHFed OpenAI models. Say goodbye to «It’s important to remember»! Prompt format is standard chatml. Don’t expect it to be good at math, riddles or be crazy smart. My end goal with AEZAKMI is to create a cozy free chatbot. Base model used for fine-tuning was 200k context Yi-6B llamafied model shared by 01.ai. I tested it up to 300k ctx. It seems to work ok up 200k. Over 200k it’s a lottery. I recommend using ChatML format, as this was used during fine-tune. Here’s a prompt format you should use, you can set a different system message, model seems to respect that fine, so it wasn’t overfitted. I recommend to set repetition penalty to something around 1.05 to avoid repetition. So far I had good experience running this model with temperature 1.2. Stories have ChatGPT like paragraph spacing, I will work on this in the future maybe, not a high priority. Unrestricted-ness of the v2 isn’t quite something that I am happy with yet, especially using prompt «A chat.». With a slightly modifed prompt it works somewhat better, I recommend…

Модальности:
Генерация текста


Задача: Генерация текста
Автор: adamo1139
Теги: llama, endpoints_compatible
Лайков: 3  |  Загрузок: 9

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.