QuantFactory/starcoder2-3b-GGUF - Каталог нейросетей
Генерация текста

QuantFactory/starcoder2-3b-GGUF

Добавлено:
QuantFactory/starcoder2-3b-GGUF

![](https://lh7-rt.googleusercontent.com/docsz/AD_4nXeiuCm7c8lEwEJuRey9kiVZsRn2W-b4pWlu3-X534V3YmVuVc2ZL-NXg2RkzSOOS2JXGHutDuyyNAUtdJI65jGTo8jT9Y99tMi4H4MqL44Uc5QKG77B0d6-JfIkZHFaUA71-RtjyYZWVIhqsNZcx8-OMaA?key=xt3VSDoCbmTY7o-cwwOFwQ) This is quantized version of bigcode/starcoder2-3b created using llama.cpp 1. Model Summary 2. Use 3. Limitations 4. Training 5. License 6. Citation StarCoder2-3B model is a 3B parameter model trained on 17 programming languages from The Stack v2, with opt-out requests excluded. The model uses Grouped Query Attention, a context window of 16,384 tokens with a sliding window attention of 4,096 tokens, and was trained using the Fill-in-the-Middle objective on 3+ trillion tokens. — Project Website: bigcode-project.org — Paper: Link — Point of Contact: contact@bigcode-project.org — Languages: 17 Programming languages The model was trained on GitHub code as well as additional selected data sources such as Arxiv and Wikipedia. As such it is not an instruction model and commands like «Write a function that computes the square root.» do not work well. Here are some examples to get started with the model. You can find a script for fine-tuning in StarCoder2’s…

Модальности:
Генерация текста

Области применения:
Генерация кода


Задача: Генерация текста
Автор: QuantFactory
Теги: gguf, code, model-index, endpoints_compatible
Лайков: 3  |  Загрузок: 960

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.