QuantFactory/llama3.1-cc-8B-GGUF

This is quantized version of nbeerbower/llama3.1-cc-8B created using llama.cpp

Original Model Card

This is an experimental finetune that formats the conversation data sequentially with the Llama 3 template.

Finetuned using an A100 on Google Colab for 3 epochs.

GGUF

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference API

Unable to determine this model’s pipeline type. Check the docs .

Base model

Finetuned

Finetuned

Quantized

this model