Nemotron-3.5-Lightning-30B-A3B-CRACK-GGUF
CRACK-abliterated NVIDIA Nemotron 3.5 Lightning 30B-A3B — GGUF quants for llama.cpp. Three quantizations (Q80 / Q4KM / Q2K)...
CRACK-abliterated NVIDIA Nemotron 3.5 Lightning 30B-A3B — GGUF quants for llama.cpp. Three quantizations (Q80 / Q4KM / Q2K)...
93.7% HarmBench compliance with only -2.0% MMLU. Full abliteration of the dense Gemma 4 31B. Tested with greedy...
Built for vMLX — the only MLX inferencer with VL support, KV cache quantization, prefix cache reuse, agentic...
Компактная модель кодирования Laguna MoE · llama.cpp GGUF · Q6K · Q4KM · Q2K Отказ удален с сохранением...