Hy3-REAP-200B-21B
A 33%-expert-pruned version of Tencent Hunyuan Hy3 (295B / A21B), compressed with REAP (Router-weighted Expert Activation Pruning) —...
A 33%-expert-pruned version of Tencent Hunyuan Hy3 (295B / A21B), compressed with REAP (Router-weighted Expert Activation Pruning) —...
This repository contains a REAM-compressed version of Tencent Hy3, produced with Akicou/ream — a REAM/REAP-style Mixture-of-Experts compression framework....
Pre-converted weights for colibrì — a pure-C engine that runs huge MoE models on consumer hardware by keeping...
The high-memory Apple Silicon edition of SuperHY3, with the same fused OBLITERATUS and SuperTune behavior as the NVFP4...
Quantized tencent/Hy3 for Apple Silicon MLX / JANG runtimes — a 295B-total / 21B-active text MoE, packed to...
Version 26.05.01 Calibration STEM and Agentic Languages EN ZH HI AR RU JA KO NL FR ES Model...
High-quality imatrix GGUF quantizations of tencent/Hy3, Tencent’s 295B-parameter Mixture-of-Experts model with ~21B active parameters per token. Produced with...
INVALID LANGUAGE PAIR SPECIFIED. EXAMPLE: LANGPAIR=EN|IT USING 2 LETTER ISO OR RFC3066 LIKE ZH-CN. ALMOST ALL LANGUAGES SUPPORTED...
Hy3, самоквантованный до GGUF с помощью Atomic Chat. Создано прямо на основе исходных весов Tencent с матрицей важности...
> Выпущено 16 июля 2026 г. Все этапы пройдены: проверка артефакта (—check-mtp), > качество по сравнению с FP8...