Edit model card

QuantFactory/Llama-3-Ko-Luxia-Instruct-GGUF

This is quantized version of maywell/Llama-3-Ko-Luxia-Instruct created using llama.cpp

Template

ChatML

Downloads last month: 667

GGUF

Model size

8.17B params

Architecture

llama

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Examples

Text Generation

This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Model tree for QuantFactory/Llama-3-Ko-Luxia-Instruct-GGUF

Base model

maywell/Llama-3-Ko-Luxia-Instruct

Quantized

(2)

this model