Llama-3.1-Nemotron-70B-Instruct-HF-FP8-DYNAMIC
— Model Architecture: Llama-3.1-Nemotron — Input: Text — Output: Text — Model Optimizations: — Weight quantization: FP8 —...
— Model Architecture: Llama-3.1-Nemotron — Input: Text — Output: Text — Model Optimizations: — Weight quantization: FP8 —...
The Meta Llama 3.2 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction-tuned...
!image/png This model was quantized by SanctumAI. To leave feedback, join our community in Discord.* Model creator: meta-llama...
The Model mlx-community/Llama-3.2-3B-8bit was converted to MLX format from meta-llama/Llama-3.2-3B using mlx-lm version 0.17.1. Модальности:Генерация текста Задача: Генерация...
The Meta Llama 3.2 collection of multilingual large language models (LLMs) is a collection of pretrained and instruction-tuned...
— Model Architecture: Meta-Llama-3.2 — Input: Text — Output: Text — Model Optimizations: — Weight quantization: FP8 —...
Now for something a bit different, VioletTwilight-v0.2! This model is a SLERP merge of AzureDusk-v0.2 and Crimson_Dawn-v0.2! The...
Romulus is a series of continually pre-trained models enriched in French law and intended to serve as the...
Contact me if you have suggestions for models to compress next or if you want to compress your...
 This is quantized version of aifeifei799/Llama-3.1-8B-Instruct-Fei-v1-Uncensored created using llama.cpp — Llama-3.1-8B-Instruct Uncensored — more informtion look at...