notux-8x7b-v1-AWQ
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
This repo contains the model checkpoints for: — model family llama13b — optimized with the loss SFT+KTO —...
NeuralHermes представляет собой модель teknium/OpenHermes-2.5-Mistral-7B, которая была дополнительно доработана с помощью оптимизации прямых предпочтений (DPO) с использованием набора...
Notus is a collection of fine-tuned models using Direct Preference Optimization (DPO) and related RLHF techniques. This model...
This model is a fine-tuned version of alignment-handbook/zephyr-7b-sft-full on the HuggingFaceH4/ultrafeedback_binarized dataset. It achieves the following results on...
This repository contains the DPO-v2 LoRA adapter for Academic Humanize, a post-training project for reducing AI-like patterns in...