LiquidAI/LFM2-1.2B-Extract finetuned on Japanese content to output JSON with PII. The model should output only single json with four fields: Evaluation on test split of stockmark/ner-wikipedia-dataset: Test accuracy on raw model using wiki dataset: 0.9100 —> 1.0 after fine-tunning**. That dataset is somehow simple. We generated 64 samples of long contracts containing PII in Japanese. We used it only to evaluate the final perfomrance of the models Test accuracy on raw model using OUR dataset: 0.5781 —> 0.9688 after fine-tunning**. We use an exact match on generated JSON. The output of SLM must be a valid JSON with exactly four required fields, no less, no more. — Developed by: @kainoj, @valeriosalvucci and @gangadhara691 — Language(s) (NLP): Japanese — License: lfm1.0 — Finetuned from model: LiquidAI/LFM2-1.2B-Extract — Finetued on dataset: stockmark/ner-wikipedia-dataset
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: kainoj
Теги: lfm2, trl, sft, conversational, ja, endpoints_compatible
Лайков: 4 | Загрузок: 17
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.