SeQwence-14B
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This...
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This...
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This...
Модель в настоящее время следует обновлению от GLM-4-9B-Chat и теперь требует трансформаторов>=4.44.0. Обновите свои зависимости соответствующим образом. Также...
— Fine-tuning of Llama-3.1-8B on german datasets. Same datasets used in Nekochu/Llama-2-13B-German-ORPO. — I’ve (alway) kept LoRA QLoRAGerman-ORPO...
Original model: https://huggingface.co/nothingiisreal/L3-8B-Celeste-V1.2 Thank you kalomaze and Dampf for assistance in creating the imatrix calibration dataset Thank you...
Эта модель представляет собой тонко настроенную версию meta-llama/Meta-Llama-3-8B-Instruct по английскому переводу небольшого набора данных из 16 000 корейских...
With a strong commitment to enhancing the quality of large language models for the Vietnamese language, a collaborative...
Merely Phase 1 UNA, only MLP’s and its kinda of a beta. The goal was to produce a...
PsychAgent-Qwen3-32B — это модель психологического консультирования, построенная на основе Qwen/Qwen3-32B. Это 32B инстанциация PsychAgent, основы обучения на протяжении...
EditScore — это серия современных моделей вознаграждений с открытым исходным кодом (7B–72B), предназначенных для оценки и улучшения редактирования...