INVALID LANGUAGE PAIR SPECIFIED. EXAMPLE: LANGPAIR=EN|IT USING 2 LETTER ISO OR RFC3066 LIKE ZH-CN. ALMOST ALL LANGUAGES SUPPORTED BUT SOME MAY HAVE NO CONTENT The PR’s converter sets supportsmtpexport = False, so the speculative > draft head is not in these files. You lose the draft-model speed-up, not any output quality. > — The GGUF format for this architecture can still change before merge.** A maintainer has asked > for the 50 GB PLE tensor to be re-sharded; if that lands, files built now (including these) may > need regenerating. Treat these as tracking an open PR, not a frozen release. No imatrix was used — these are plain K-quants, which do not need one. An importance-matrix run means inference over the full 330 GB model, and an i-quant (IQ*) without one is the weakest option in the list, so none is shipped rather than ship it unmarked. If you want IQ quants with a proper imatrix, the BF16 file is here…
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: windowsxp811203
Теги: gguf, llama.cpp, abliterated, uncensored, qwen4_exp, qwen3.8, flash-next, moe
Лайков: 4 | Загрузок: 127,983
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.