This model is the quantized version of NexusFlow’s NexusRaven V-2. The quantization technique used is Activation-Aware Weight Quantization. The model is suitable for high-degree of Function Calling. The functions may be Simple Functions, Compound Functions or Nested Functions. The model hasn’t been fine-tuned yet.
Модальности:
Генерация текста
Задача: Генерация текста
Автор: NaiveAttention
Теги: llama, function calling, text-generation-inference, endpoints_compatible, 4-bit, awq
Лайков: 3 | Загрузок: 20
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.