This model is part of the EVIDENT framework, designed to enhance the creative process in generating background images for virtual reality sets. It interprets user instructions to generate and modify prompts for text-to-image models. This is the KTO version of the model, you can also check at the SFT and DPO versions. The demo integrates a diffusion model to test prompt-image alignment, and mechanisms for user feedback and iterative prompt refinement, aiming to enhance user creativity and satisfaction. The instruction categories are: — Addition: Involves the inclusion of new elements or features. — Condensation: Consists in the summarization of the description. — Modification: Alters specific aspects of the description to change the scene. — Rearrangement: Reordering of sentences within the descriptions. — Removal: Elimination of specific details in the description. — Rephrase: Rewriting parts of the description. — Scene Change: Overall description context switch. The output language of the model is English, but other languages can be used as input (quality depends of the quantity of tokens used on the pre-training phase for the given language). Developed as part of the EVIDENT…
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: ITG
Теги: mistral, chatml, synthetic data, finetune, kto, conversational, en, text-generation-inference
Лайков: 3 | Загрузок: 0
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.