End of training

Files changed (4) hide show

README.md ADDED Viewed

+---
+license: apache-2.0
+base_model: distilroberta-base
+tags:
+- generated_from_trainer
+model-index:
+- name: distilroberta-base-finetuned-wikitext2
+  results: []
+---
+<!-- This model card has been generated automatically according to the information the Trainer had access to. You
+should probably proofread and complete it, then remove this comment. -->
+# distilroberta-base-finetuned-wikitext2
+This model is a fine-tuned version of [distilroberta-base](https://huggingface.co/distilroberta-base) on the None dataset.
+It achieves the following results on the evaluation set:
+- Loss: 1.8384
+## Model description
+More information needed
+## Intended uses & limitations
+More information needed
+## Training and evaluation data
+More information needed
+## Training procedure
+### Training hyperparameters
+The following hyperparameters were used during training:
+- learning_rate: 2e-05
+- train_batch_size: 8
+- eval_batch_size: 8
+- seed: 42
+- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
+- lr_scheduler_type: linear
+- num_epochs: 3.0
+### Training results
+| Training Loss | Epoch | Step | Validation Loss |
+|:-------------:|:-----:|:----:|:---------------:|
+| 2.0836        | 1.0   | 2403 | 1.8844          |
+| 1.985         | 2.0   | 4806 | 1.8668          |
+| 1.952         | 3.0   | 7209 | 1.8682          |
+### Framework versions
+- Transformers 4.35.2
+- Pytorch 2.1.0+cu121
+- Datasets 2.16.1
+- Tokenizers 0.15.0

model.safetensors CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:4f27acdb337e0a17c410650d19a2ca0d4f3f3ca76b40fd047635dca7b06a6d6b
 size 328693404

 version https://git-lfs.github.com/spec/v1
+oid sha256:23e1e4f8fdf5889208168791888a9d19948f214fac00cf6dae0fda16834fc02e
 size 328693404

runs/Jan20_12-47-52_90d56b2707cd/events.out.tfevents.1705754872.90d56b2707cd.5917.3 CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:7f65c246e118dc49b91b03d54679058a4fbf1a652a8274962d2db908280bb391
-size 7022

 version https://git-lfs.github.com/spec/v1
+oid sha256:2e99c5fc67a5a254aa21a7082540768d44248a055fe1c5a3705ab934baf94cea
+size 7647

runs/Jan20_12-47-52_90d56b2707cd/events.out.tfevents.1705755329.90d56b2707cd.5917.4 ADDED Viewed

+version https://git-lfs.github.com/spec/v1
+oid sha256:bd61382f45a14ae2f78ecb38ee01dd15e227189a2857cfbd30a07b1e97352bff
+size 359