XtraGPT is a family of open-source Large Language Models (LLMs) designed specifically for human-AI collaborative academic paper revision. Unlike general-purpose models that often perform surface-level polishing, XtraGPT is fine-tuned to understand the full context of a research paper and execute specific, criteria-guided revision instructions. The models were trained on a dataset of 140,000 high-quality instruction-revision pairs derived from top-tier conference papers (ICLR). Key Features: Context-Aware: Processes the full paper context to ensure revisions maintain consistency with the global narrative. Controllable: Follows specific user instructions aligned with 20 academic writing criteria across 6 sections (Abstract, Introduction, etc.). Iterative Workflow:** Designed to support the «Human-AI Collaborative» (HAC) lifecycle where authors retain creative control. Available Model Sizes: 1.5B (Based on Qwen/Qwen2.5-1.5B-Instruct) 3B (Based on meta-llama/Llama-3.2-3B-Instruct) 7B (Based on Qwen/Qwen2.5-7B-Instruct) 14B (Based on microsoft/phi-4) XtraGPT is compatible with vLLM for high-throughput inference.
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: Xtra-Computing
Теги: llama, chat, conversational, zho, eng, fra, spa, por
Лайков: 3 | Загрузок: 38
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.