This is quantized version of bigscience/bloom-7b1 created using llama.cpp BLOOM LM BigScience Large Open-science Open-access Multilingual Language Model Model Card 1. Model Details 2. Uses 3. Training Data 4. Risks and Limitations 5. Evaluation 6. Recommendations 7. Glossary and Calculations 8. More Information 9. Model Card Authors This section provides information for anyone who wants to know about the model. All collaborators are either volunteers or have an agreement with their employer. (Further breakdown of participants forthcoming.)* Cite as: BigScience, BigScience Language Open-science Open-access Multilingual (BLOOM) Language Model. International, May 2021-May 2022 Organizations of contributors. (Further breakdown of organizations forthcoming.)* This section provides information for people who work on model development. Please see the BLOOM training README for full details on replicating training. Model Architecture: Modified from Megatron-LM GPT2 (see paper, BLOOM Megatron code): Layer normalization applied to word embeddings layer (StableEmbedding`; see code, paper) * ALiBI positional encodings (see paper), with GeLU activation functions * Sequence length of 2048…
Модальности:
Генерация текста
Области применения:
Генерация кода
Задача: Генерация текста
Автор: QuantFactory
Теги: gguf, ak, ar, as, bm, bn, ca, code
Лайков: 3 | Загрузок: 1,072
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.