MobileMoE-S-QAT
MobileMoE — это семейство языковых моделей Mixture-of-Experts (MoE) на устройстве с субмиллиардными активными параметрами, предназначенных для продвижения границы...
MobileMoE — это семейство языковых моделей Mixture-of-Experts (MoE) на устройстве с субмиллиардными активными параметрами, предназначенных для продвижения границы...
LFM2-12B-A1B-GLM-4.7-Thinking-Quantum-IQ1C-P-TR1-S2 [Stage 2] Fine tune of «LFM2-8B-A1B-GLM-4.7-Flash-Thinking-Quantum-IQ1C-P» using Unsloth using custom dataset(s), 128k context in 16 bit precision....
Fine tune of «LFM2-8B-A1B» using Unsloth using custom dataset(s), 128k context in 16 bit precision. This model is...
Верная мелкомасштабная (~ 8,1B всего / ~ 2,21B активируется на токен) реплика архитектуры DeepSeek-V4, рассчитанная на обучение на...
Соответствующая архитектуре стандартная базовая линия МОС, выпущенная вместе с EMO: Смесь экспертов по предварительной подготовке для возникающей модульности...
static quants of https://huggingface.co/lordx64/Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled For a convenient overview and download list, visit our model page for this model....
GGUF quantizations of DJLougen/Ornstein3.6-35B-A3B-SABER for use with llama.cpp, ollama, LM Studio, and compatible runtimes. Source model is the...
This is an MLX release of an abliterated version of Qwen’s Qwen3.6-35B-A3B. By applying Heretic’s ablation pipeline to...
This is the 3-bit Apple MLX release of an abliterated version of MiniMaxAI’s MiniMax-M2.7. By applying Heretic’s Ablated...
Квантование MLX nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16 для Apple Silicon. — Apple Silicon Mac с унифицированной памятью 128 ГБ — mlx-lm >=...