h2ogpt-4096-llama2-7b
Эта модель может быть точно настроена с помощью программного обеспечения с открытым исходным кодом H2O.ai: — h2oGPT https://github.com/h2oai/h2ogpt/...
Эта модель может быть точно настроена с помощью программного обеспечения с открытым исходным кодом H2O.ai: — h2oGPT https://github.com/h2oai/h2ogpt/...
This is a sharded version of Meta’s Llama 2 chat 7b model, specifically the hugging face version. Shards...
The weight file is split into chunks with a size of 405MB for convenient and fast parallel downloads...
— BaseModel: Meta’s Llama 2 7B chat hf. — Dataset: timdettmers/openassistant-guanaco. We are unlocking the power of large...
Speedup inference while reducing memory by 2x-4x using int8 inference in C++ on CPU or GPU. Checkpoint compatible...
MobileMoE — это семейство языковых моделей Mixture-of-Experts (MoE) на устройстве с субмиллиардными активными параметрами, предназначенных для продвижения границы...
We are introducing MobileLLM-P1 or Pro, a 1B foundational language model in the MobileLLM series, designed to deliver...
The NeuraLake iSA-03-Mini-3B (Hybrid) is an advanced AI model developed by NeuraLake, specifically designed to integrate the best...
Наш последний метод квантования вводит прецизионно-адаптивное квантование для сверхнизкоразрядных моделей (1-2 бита) с проверенными улучшениями на Llama-3-8B. Этот...
This model was generated using llama.cpp at commit f5cd27b7. Our latest quantization method introduces precision-adaptive quantization for ultra-low-bit...