Speck1-140M-Instruct
Speck1-140M-Instruct is a 140.7M parameter English instruction-tuned language model. It was initialized from Speck1-140M, a hybrid model that...
Speck1-140M-Instruct is a 140.7M parameter English instruction-tuned language model. It was initialized from Speck1-140M, a hybrid model that...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
Partial GPTQ Int4 quant of Qwen/Qwen3.6-27B, produced with the verbatim recipe from Qwen’s own Qwen/Qwen3.5-27B-GPTQ-Int4 — only MLP/FFN...
GKA-primed-HQwen3-8B-Reasoner — это гибридная языковая модель, состоящая из 50% слоев внимания и 50% слоев стробированного KalmaNet (GKA), загрунтованных...
RWKV-GLM-4.7-Flash-exp is an alpha-stage experimental model that converts GLM-4.7-Flash into a fully linear-attention-dominant architecture using the RADLADS distillation...
> ⚠️ CONTENT WARNING: RATED R — MATURE AUDIENCES ONLY > > This AI model is a form...
This is a hybrid Mamba-Transformer model based on the Llama 3.1 architecture, distilled from Llama 3.3 70B into...
Это гибридная модель Mamba-Transformer, основанная на архитектуре Llama 3.2, перегоненная из Llama 3.1 8B в модель параметров 3B...
☕ Если эти модели вам полезны, рассмотрите возможность поддержки моей работы — это позволит вычислить больше и больше...