Lyra4-Gutenberg-12B / README.md
Sao10K's picture
Update README.md
4d77a82 verified
|
raw
history blame
518 Bytes
metadata
license: cc-by-nc-4.0
library_name: transformers
base_model:
  - Sao10K/MN-12B-Lyra-v4
datasets:
  - jondurbin/gutenberg-dpo-v0.1

Lyra4-Gutenberg-12B

Sao10K/MN-12B-Lyra-v4 finetuned on jondurbin/gutenberg-dpo-v0.1.

Method

ORPO Finetuned using an RTX 3090 + 4060 Ti for 3 epochs.

Fine-tune Llama 3 with ORPO