ibrahimkettaneh/QwQ-32B-Preview-abliterated-4.5bpw-h8-exl2 - Каталог нейросетей
Генерация текста

ibrahimkettaneh/QwQ-32B-Preview-abliterated-4.5bpw-h8-exl2

Добавлено:
ibrahimkettaneh/QwQ-32B-Preview-abliterated-4.5bpw-h8-exl2

Source: 🐺🐦‍⬛ LLM Comparison/Test: 25 SOTA LLMs (including QwQ) through 59 MMLU-Pro CS benchmark runs Credits go to for their helpful and informative benchmark: Wolfram Ravenwolf To increase performance, increase the max new output when running inference from the default to 16384 tokens. For more context, details, and comparisons, you can refer to the original article by Ravenwolf. This is an uncensored version of Qwen/QwQ-32B-Preview created with abliteration (see remove-refusals-with-transformers to know more about it). This is a crude, proof-of-concept implementation to remove refusals from an LLM model without using TransformerLens. QwQ-32B-Preview is an experimental research model developed by the Qwen Team, focused on advancing AI reasoning capabilities. As a preview release, it demonstrates promising analytical abilities while having several important limitations: 1. Language Mixing and Code-Switching: The model may mix languages or switch between them unexpectedly, affecting response clarity. 2. Recursive Reasoning Loops: The model may enter circular reasoning patterns, leading to lengthy responses without a conclusive answer. 3. Safety and Ethical Considerations: The model…

Модальности:
Генерация текста

Области применения:
Диалог / чат


Задача: Генерация текста
Автор: ibrahimkettaneh
Теги: qwen2, chat, abliterated, uncensored, conversational, en, text-generation-inference, endpoints_compatible
Лайков: 4  |  Загрузок: 12

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.