Magnolia-v3b-12B-GGUF

This repo is a set of GGUF quants of a grimjim/Magnolia-v3b-12B, a merge of pre-trained language models created using mergekit.

llama.cpp was used to make the following quants:

Downloads last month
1
GGUF
Model size
12.2B params
Architecture
llama
Hardware compatibility
Log In to view the estimation

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for grimjim/Magnolia-v3b-12B-GGUF

Quantized
(4)
this model