— Base Model: IntelLabs/sqft-phi-3-mini-4k-50-base — Sparsity: 50% — Quantization: No — Finetune Method: SQFT + SparsePEFT — Finetune data: 10K instruction-following math reasoning training dataset from LLM-Adapters (math_10k.json) — Sub-Adapter: Heuristic Refer to our repo for the environment information to run this command. Repository: https://github.com/IntelLabs/Hardware-Aware-Automated-Machine-Learning/tree/main/SQFT Paper: — SQFT: Low-cost Model Adaptation in Low-precision Sparse Foundation Models — Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
Модальности:
Генерация текста
Области применения:
Диалог / чат Математика
Задача: Генерация текста
Автор: IntelLabs
Теги: phi3, conversational, custom_code, en, text-generation-inference, endpoints_compatible
Лайков: 3 | Загрузок: 24
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.