OpenGVLab/ASMv2-Stage2-Pretrain - Каталог нейросетей
Генерация текста

OpenGVLab/ASMv2-Stage2-Pretrain

Добавлено:
OpenGVLab/ASMv2-Stage2-Pretrain

This is a pretrained checkpoint, you can use it to instruct tune your multimodal models. Model type: ASMv2 is an open-source chatbot trained by fine-tuning LLaMA/Vicuna on multimodal instruction-following data. It integrates the Relation Conversation (ReC) ability while maintaining powerful general capabilities. This model is also endowed with grounding and referring capabilities, exhibiting state-of-the-art performance on region-level tasks, and can be naturally adapted to the Scene Graph Generation task in an open-ended manner. Paper or resources for more information: https://github.com/OpenGVLab/all-seeing ASMv2-Pretrain is open-sourced under the Apache License 2.0, Where to send questions or comments about the model: https://github.com/OpenGVLab/all-seeing/issues Primary intended uses: The primary use of ASMv2 is research on large multimodal models and chatbots. Primary intended users: The primary intended users of the model are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence. The pretrain phase employs 5M filtered samples from CC12M, 10M filtered samples from AS-1B, and 15M filtered samples from GRiT.

Модальности:
Генерация текста


Задача: Генерация текста
Автор: OpenGVLab
Теги: llava, endpoints_compatible
Лайков: 3  |  Загрузок: 34

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.