Magnolia-v3b-12B-GGUF

This repo is a set of GGUF quants of a grimjim/Magnolia-v3b-12B, a merge of pre-trained language models created using mergekit.

llama.cpp was used to make the following quants:

Downloads last month
128
GGUF
Model size
12.2B params
Architecture
llama
Hardware compatibility
Log In to view the estimation

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for grimjim/Magnolia-v3b-12B-GGUF

Quantized
(2)
this model