This model was converted to GGUF format from duoqi/Nanbeige2-16B-Chat using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to the original model card for more details on the model. Step 2: Move into the llama.cpp folder and build it with LLAMACURL=1 flag along with other hardware-specific flags (for ex: LLAMACUDA=1 for Nvidia GPUs on Linux).
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: NikolayKozloff
Теги: gguf, llm, llama-cpp, gguf-my-repo, en, zh, endpoints_compatible, conversational
Лайков: 3 | Загрузок: 13
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.