yzhuang commited on
Commit
7de45f8
1 Parent(s): da1566d

End of training

Browse files
README.md CHANGED
@@ -15,7 +15,7 @@ model-index:
15
  <!-- This model card has been generated automatically according to the information the Trainer had access to. You
16
  should probably proofread and complete it, then remove this comment. -->
17
 
18
- [<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/yufanz/autotree/runs/7283910327.75521-df0dd9e4-b029-4f7b-b0df-488a352215cc)
19
  # Meta-Llama-3-8B-Instruct_fictional_arc_challenge_Korean_v2
20
 
21
  This model is a fine-tuned version of [meta-llama/Meta-Llama-3-8B-Instruct](https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct) on the generator dataset.
@@ -45,7 +45,7 @@ The following hyperparameters were used during training:
45
  - total_train_batch_size: 16
46
  - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
47
  - lr_scheduler_type: linear
48
- - num_epochs: 36
49
 
50
  ### Training results
51
 
 
15
  <!-- This model card has been generated automatically according to the information the Trainer had access to. You
16
  should probably proofread and complete it, then remove this comment. -->
17
 
18
+ [<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/yufanz/autotree/runs/7283881766.46478-720696f5-5799-4c26-9505-2df28e3a300e)
19
  # Meta-Llama-3-8B-Instruct_fictional_arc_challenge_Korean_v2
20
 
21
  This model is a fine-tuned version of [meta-llama/Meta-Llama-3-8B-Instruct](https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct) on the generator dataset.
 
45
  - total_train_batch_size: 16
46
  - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
47
  - lr_scheduler_type: linear
48
+ - num_epochs: 48
49
 
50
  ### Training results
51
 
model-00001-of-00004.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:d3ce68b1716618b77f0ad0cc9c777b20a913b4732689192b47b4e464d823729d
3
  size 4976698672
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:39d5f06f5a6b492845b09b91f6a70e641c1ca4e3bf3d877a041baa5200020468
3
  size 4976698672
model-00002-of-00004.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:3226eb99f36321e996d3016512e2b6fb790312c06be794db7d05ec2b8e56bb89
3
  size 4999802720
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:ee6fb10385f5b8e1c9b962707b1b7c9186a02b3b3978507cd0c7fc4ef9bca680
3
  size 4999802720
model-00003-of-00004.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:0566aca7979b3e87ee4c1bbdc1f67989a0cc03d0a2cf29f5760fe57ba4f06e78
3
  size 4915916176
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:dfb49419bfef52a81031053121715dcef21340b47c22da974db8cadd09c61ffd
3
  size 4915916176
model-00004-of-00004.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:ebc77d24e79040530f6c2a879b7c5ebf7d7db34ec5c2fbd8adfd80290a73676f
3
  size 1168138808
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:75c1fc26f72b6ac4c48cfe0c89554ad89b3401dfb7e5fe93fb8d2a4490a99e80
3
  size 1168138808
runs/May19_04-38-43_node-0/events.out.tfevents.1716093526.node-0.4095.0 ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:d3cf0f0bd848eb391eb050d9e6cc4dd04ec1db6240ed2e0d33c57183cbb65462
3
+ size 5289
training_args.bin CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:0d192babb9de57d3473be92bb496c62dce3ec847dca34fa4bcf90360afb4171f
3
  size 5176
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:22f7e62ae8c681ee4462d466d9cc727f7d9ccf85c4a163f6a8b4f32b3c344839
3
  size 5176