Llama-3.2-3B-CodeReactor
This is a merge of pre-trained language models created using mergekit. This model was merged using the passthrough...
This is a merge of pre-trained language models created using mergekit. This model was merged using the passthrough...
Unleashing Reasoning Capability of LLMs via Scalable Question Synthesis from Scratch We introduce ScaleQuest, a scalable and novel...
This is a merge of pre-trained language models created using mergekit. This model was merged using the passthrough...
This is a merge of pre-trained language models created using mergekit. This model was merged using the Model...
The Llama3-8B-1.58 models are large language models fine-tuned on the BitNet 1.58b architecture, starting from the base model...
This language model is a merged version of several pre-trained models, designed to excel in roleplay, long-form question...
This is a merge of pre-trained language models created using mergekit. This model was merged using the Model...
This model is a distilled version of LLaMA 2, containing approximately 80 million parameters. It was trained using...
We introduce PRefLexOR (Preference-based Recursive Language Modeling for Exploratory Optimization of Reasoning), a framework that combines preference optimization...
This is a merge of pre-trained language models created using mergekit. The following models were included in the...