SjardiWillems commited on
Commit
44b6d33
·
verified ·
1 Parent(s): 831f04d

End of training

Browse files
README.md ADDED
@@ -0,0 +1,65 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ base_model: distilbert-base-uncased
4
+ tags:
5
+ - generated_from_trainer
6
+ metrics:
7
+ - spearmanr
8
+ model-index:
9
+ - name: distilbert-base-uncased-finetuned-stsb
10
+ results: []
11
+ ---
12
+
13
+ <!-- This model card has been generated automatically according to the information the Trainer had access to. You
14
+ should probably proofread and complete it, then remove this comment. -->
15
+
16
+ # distilbert-base-uncased-finetuned-stsb
17
+
18
+ This model is a fine-tuned version of [distilbert-base-uncased](https://huggingface.co/distilbert-base-uncased) on an unknown dataset.
19
+ It achieves the following results on the evaluation set:
20
+ - Loss: 0.5575
21
+ - Pearson: 0.8689
22
+ - Spearmanr: 0.8661
23
+
24
+ ## Model description
25
+
26
+ More information needed
27
+
28
+ ## Intended uses & limitations
29
+
30
+ More information needed
31
+
32
+ ## Training and evaluation data
33
+
34
+ More information needed
35
+
36
+ ## Training procedure
37
+
38
+ ### Training hyperparameters
39
+
40
+ The following hyperparameters were used during training:
41
+ - learning_rate: 2e-05
42
+ - train_batch_size: 16
43
+ - eval_batch_size: 16
44
+ - seed: 42
45
+ - optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
46
+ - lr_scheduler_type: linear
47
+ - num_epochs: 5
48
+
49
+ ### Training results
50
+
51
+ | Training Loss | Epoch | Step | Validation Loss | Pearson | Spearmanr |
52
+ |:-------------:|:-----:|:----:|:---------------:|:-------:|:---------:|
53
+ | No log | 1.0 | 360 | 0.6899 | 0.8641 | 0.8626 |
54
+ | 1.0355 | 2.0 | 720 | 0.6125 | 0.8668 | 0.8647 |
55
+ | 0.393 | 3.0 | 1080 | 0.5795 | 0.8688 | 0.8655 |
56
+ | 0.393 | 4.0 | 1440 | 0.5575 | 0.8689 | 0.8661 |
57
+ | 0.2366 | 5.0 | 1800 | 0.5678 | 0.8679 | 0.8649 |
58
+
59
+
60
+ ### Framework versions
61
+
62
+ - Transformers 4.38.2
63
+ - Pytorch 2.1.0+cu121
64
+ - Datasets 2.18.0
65
+ - Tokenizers 0.15.2
runs/Mar05_21-56-47_fe13c9557656/events.out.tfevents.1709675807.fe13c9557656.828.0 CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:57ba0bb28d2969ea99c739f7ac236f7d7f6785be14cff4bf43b2249ba774f1c8
3
- size 6513
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:eedba4601221130254cdebcd4fd32c581bfb1bd46c85f934eaf495620feb40fc
3
+ size 7453
runs/Mar05_21-56-47_fe13c9557656/events.out.tfevents.1709675899.fe13c9557656.828.1 ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:9624145633ba745822aae65f4c199953be61bb4d61849a1f7c97d8775acffa5b
3
+ size 463