Qwen 3 Embedding 8B in INT4. About 5 GB on disk. Runs on an 8 GB consumer GPU. ~6.1 GB on disk. Recommended VRAM: enough headroom for KV cache. This artifact is a derivative work of Qwen/Qwen3-Embedding-8B, released by its original authors under the Apache License, Version 2.0. This artifact is distributed under the same license. The full license text is included in LICENSE, and required attribution is in NOTICE. License text: https://www.apache.org/licenses/LICENSE-2.0 Source model: https://huggingface.co/Qwen/Qwen3-Embedding-8B
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: drawais
Теги: qwen3, quantized, 4-bit, int4, awq, conversational, en, text-generation-inference
Лайков: 4 | Загрузок: 4,328
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.