LongWriter-glm4-9b-abliterated-gguf
This model was converted to GGUF format from byroneverson/LongWriter-glm4-9b-abliterated using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to...
This model was converted to GGUF format from byroneverson/LongWriter-glm4-9b-abliterated using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to...
This is quantized version of THUDM/codegeex4-all-9b created using llama.cpp We introduce CodeGeeX4-ALL-9B, the open-source version of the latest...
> Use this image: September 9 reference stack > verdictai/trellismx:glm53-flash-p8-r27-reference-20260909 > This is the selected serving image for...
INVALID LANGUAGE PAIR SPECIFIED. EXAMPLE: LANGPAIR=EN|IT USING 2 LETTER ISO OR RFC3066 LIKE ZH-CN. ALMOST ALL LANGUAGES SUPPORTED...
Spark is a quality-first 2.80 BPW GGUF quant of zai-org/GLM-5.3-Flash, tuned to fit and run fully on a...
GGUF quantizations of zai-org/GLM-5.3-Flash, made with llama.cpp. 320B total parameters, 18B active. 45 layers with hybrid attention: 34...
This is a full-expert GLM-5.2 hybrid checkpoint that combines the compact MXFP8/NVFP4/NF3 layout from madeby561/GLM-5.2-MXFP8-NVFP4-NF3-Hybrid with selected refusal-reduction...
Non-uniform expert prune · NVFP4 · runs on FOUR Blackwell cards · verified 0-refusal Three transformations in one...
Массы квантуют до INT4 (размер группы 128); активации выполняют в BF16. Результатом является контрольная точка объемом 388 ГБ...
Эта работа относится к установленной области исследований интерпретируемости и безопасности LLM. Аблитерация — это задокументированная методика изучения механизмов...