Llama-3-Orca-1.0-8B-GGUF

Quant of https://huggingface.co/Locutusque/Llama-3-Orca-1.0-8B

f32
f16
Q8_0
Q4_K_M
Q2_K

Downloads last month: 1,155

GGUF

Model size

8.03B params

Architecture

llama

2-bit

4-bit

5-bit

8-bit

16-bit

Inference API

Text Generation

This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Collection including leafspark/Llama-3-Orca-1.0-8B-GGUF

Llama 3 Experiments

Collection

3 items • Updated May 20