YuYu1015/YuYu1015-Ornith-1.0-9B-abliterated-dpo - Каталог нейросетей
Генерация текста

YuYu1015/YuYu1015-Ornith-1.0-9B-abliterated-dpo

Добавлено:
YuYu1015/YuYu1015-Ornith-1.0-9B-abliterated-dpo

> | repeat-penalty 1.05 ✅ | correct (sweet spot) | > | repeat-penalty 1.0 | severe thinking loops | > | repeat-penalty 1.1 | truncated / unfinished answers | > | temp 0 (greedy) | not recommended | An abliterated (uncensored) variant of deepreinforce-ai/Ornith-1.0-9B, a Qwen3.5-architecture reasoning model. Refusal behavior has been removed and moralizing/disclaimer tendencies substantially reduced, while preserving — and slightly improving — reasoning ability. Measured on 200 harmful-intent prompts (refusal / moralizing) and GSM8K (reasoning). Refusal and moralizing are detected with independent BERT classifiers; GSM8K is exact-match accuracy. → Refusals essentially eliminated, moralizing cut to roughly one third, and reasoning not degraded (in fact slightly higher). This is a reasoning model — keep thinking enabled and use the official Qwen3.5 sampling settings: > ⚠️ —repeat-penalty is critical — keep it at 1.05. This value gives near-normal generation and is the sweet spot for this model. Do NOT change it: 1.0 causes severe thinking loops, while 1.1 makes the model fail to finish its answer. Greedy decoding (—temp 0`) is also not recommended for this family. This model has…

Модальности:
Генерация текста

Области применения:
Логика и рассуждение Диалог / чат


Задача: Генерация текста
Автор: YuYu1015
Теги: qwen3_5, image-text-to-text, qwen3.5, gated-deltanet, reasoning, thinking, abliterated, uncensored
Лайков: 4  |  Загрузок: 48

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.