CroissantLLMBase-GGUF
This model is part of the CroissantLLM initiative, and corresponds to the checkpoint after 190k steps (2.99 T)...
This model is part of the CroissantLLM initiative, and corresponds to the checkpoint after 190k steps (2.99 T)...
This model was converted to MLX format from [codellama/CodeLlama-13b-Instruct-hf](). Refer to the original model card for more details...
This model was converted to MLX format from [codellama/CodeLlama-7b-Instruct-hf](). Refer to the original model card for more details...
Code Llama is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion...
Code Llama is a collection of pretrained and fine-tuned generative text models ranging in scale from 7 billion...
Model aims to generate a regex pattern given a sample string. For now, it’s only trained in non-standard...
phixtral-3x2_8 is the first Mixure of Experts (MoE) made with two microsoft/phi-2 models, inspired by the mistralai/Mixtral-8x7B-v0.1 architecture....
This modelcard is for tinymistral-v2-pycoder-instruct, a python-specific code generation model on top of Locutusque/TinyMistral-248M-v2-Instruct. This instruct model follows...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Use the Spider and BIRDSQL datasets to fine-tune deepseek-ai/deepseek-coder-6.7b in order to improve the model’s text to SQL...