This is quantrized version of HiroseKoichi/L3-8B-Lunar-Stheno created using llama.cpp L3-8B-Lunaris-v1 is definitely a significant improvement over L3-8B-Stheno-v3.2 in terms of situational awareness and prose, but it’s not without issues: the response length can sometimes be very long, causing it to go on a rant; it tends to not take direct action, saying that it will do something but never actually doing it; and its performance outside of roleplay took a hit. This merge fixes all of those issues, and I’m genuinely impressed with the results. While I did use a SLERP merge to create this model, there was no blending of the models; all I did was replace L3-8B-Stheno-v3.2’s weights with L3-8B-Lunaris-v1’s. — License: llama3 — Instruct Format: llama-3 — Context Size: 8K
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: QuantFactory
Теги: gguf, nsfw, not-for-all-audiences, llama-3, text-generation-inference, mergekit, merge, endpoints_compatible
Лайков: 3 | Загрузок: 429
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.