Edit model card

gemma2-gutenberg-9B

UCLA-AGI/Gemma-2-9B-It-SPPO-Iter3 finetuned on jondurbin/gutenberg-dpo-v0.1.

Method

Finetuned using an RTX 4090 using ORPO for 3 epochs.

Fine-tune Llama 3 with ORPO

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 22.61
IFEval (0-Shot) 27.96
BBH (3-Shot) 42.36
MATH Lvl 5 (4-Shot) 1.44
GPQA (0-shot) 11.74
MuSR (0-shot) 16.71
MMLU-PRO (5-shot) 35.47
Downloads last month
192
Safetensors
Model size
9.24B params
Tensor type
BF16
·
Inference API
This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Model tree for nbeerbower/gemma2-gutenberg-9B

Finetuned
this model
Merges
2 models
Quantizations
3 models

Dataset used to train nbeerbower/gemma2-gutenberg-9B

Evaluation results