After banging my head against the wall some more — I actually managed to merge DeepSeek distill into my mess! Along with even more models (my hand just slipped, I swear) The prose is better than in v0.5, but has a different feel to it, so I guess it’s more of a step to the side than forward (hence the title EXTRA instead of 0.6). The context recall may have improved, or I’m just gaslighting myself to think so. They kinda work out of the box if you add to the ‘Start Reply With’ field in ST — that way the model will write a really short character thought in it. However, if we want some OOC reasoning, things get trickier. My initial thought was that this model could be instructed to use either only for {{char}}’s inner monologue or for detached analysis, but actually it would end up writing character thoughts most of the time anyway, and the times when it did reason stuff it threw the narrative out of the window by making it too formal and even adding some notes at the end. And so the solution was to add a prefill after the tag. There’s a lot of room for improvement, but for now, I think this boats the float or whatever: If you add the line break after the tag, the output becomes…
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: Nohobby
Теги: llama, mergekit, merge, conversational, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 30
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.