WeDLM is an 8B parameter instruction-tuned model by Tencent, supporting English and Chinese. It features QK Norm architecture similar to Qwen3. This GGUF uses qwen3 architecture identifier for maximum llama.cpp compatibility. — Parameters: 8.19B — Layers: 36 — Hidden Size: 4096 — Attention Heads: 32 (8 KV heads, GQA) — Context Length: 16384 — Features: QK Norm, SwiGLU, RoPE (theta=1M) — Original model: Tencent WeDLM Team — Inference framework: llama.cpp This is an unofficial quantization. For official support, please refer to the original model repository.
Модальности:
Генерация текста
Области применения:
Диалог / чат Следование инструкциям
Задача: Генерация текста
Автор: feedseawave
Теги: gguf, llama-cpp, wedlm, tencent, qwen3, quantized, en, zh
Лайков: 4 | Загрузок: 58
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.