Ornith-1.5-35B-A3B-OptiQ-4bit-REAP-19B
> Built with mlx-optiq, the MLX-native toolkit to quantize, prune, fine-tune, and serve LLMs locally on Apple Silicon....
> Built with mlx-optiq, the MLX-native toolkit to quantize, prune, fine-tune, and serve LLMs locally on Apple Silicon....
MLX (mlx-vlm tree) 8-bit build of GLM-5.3-Flash (320B-A18B, glm5next), converted by streaming dequant of the official FP8 release...
An AXQuant (AXQ) mixed-precision MLX checkpoint for Apple Silicon, converted directly from the BF16 source model. The language...
SSD-backed Flash-MoE package for Qwen3.8-2.4T-A95B, exported from the Unsloth UD-Q1_0 GGUF. This is not a conventional standalone GGUF....
MLX 4-bit affine quantization of poolside/Laguna-S-2.1, packaged for Apple Silicon experiments and local OpenAI-compatible serving. Laguna S 2.1...
Quantized tencent/Hy3 for Apple Silicon MLX / JANG runtimes — a 295B-total / 21B-active text MoE, packed to...
QLoRA fine-tune of NVIDIA Nemotron-3-Super-120B-A12B (120B-total / 12B-active hybrid Mamba-2 + Latent-MoE, nemotronh) on the piagent split of...
Первая сборка MTPLX Qwopus3.6-27B-Coder — тонкая настройка агентного кодирования Jackrong Qwopus3.6-27B-v2 (репо-уровневое кодирование, многооборотная оркестровка инструмента, 67,0% SWE-bench...
This is an MTPLX pair bundle for Gemma 4 31B speculative decoding on Apple Silicon. It is not...
This is an MTPLX pair bundle for Gemma 4 31B speculative decoding on Apple Silicon. It is not...