Edit model card

QuantFactory/mistral-nemo-cc-12B-GGUF

This is quantized version of nbeerbower/mistral-nemo-cc-12B created using llama.cpp

Original Model Card

This is an experimental finetune that formats the conversation data sequentially with ChatML.

Finetuned using an A100 on Google Colab for 3 epochs.

GGUF

Model size

12.2B params

Architecture

llama

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference API

Unable to determine this model’s pipeline type. Check the docs .

Base model

Finetuned

Quantized

(4)

this model