dusersad12 commited on
Commit
8ef5ecf
·
verified ·
1 Parent(s): 64e2dc8

Upload FineTunedBest model (run_gamma, best by eval_accuracy) with filled-in benchmark scores

Browse files
Files changed (6) hide show
  1. README.md +49 -0
  2. config.json +8 -0
  3. figures/fig1.png +3 -0
  4. figures/fig2.png +3 -0
  5. figures/fig3.png +3 -0
  6. pytorch_model.bin +3 -0
README.md ADDED
@@ -0,0 +1,49 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: apache-2.0
3
+ library_name: transformers
4
+ ---
5
+ # FineTunedBest
6
+
7
+ <div align="center">
8
+ <img src="figures/fig1.png" width="65%" alt="FineTunedBest Architecture" />
9
+ </div>
10
+
11
+ ## Overview
12
+
13
+ FineTunedBest is a fine-tuned RoBERTa model optimized for multi-task performance across reasoning, comprehension, and generation benchmarks. This model was selected from a series of hyperparameter sweeps as the run with the highest evaluation accuracy.
14
+
15
+ <div align="center">
16
+ <img width="75%" src="figures/fig2.png">
17
+ </div>
18
+
19
+ ## Benchmark Results
20
+
21
+ | Benchmark | FineTunedBest |
22
+ |---|---|
23
+ | Math Reasoning | 0.550 |
24
+ | Logical Reasoning | 0.819 |
25
+ | Reading Comprehension | 0.700 |
26
+ | Code Generation | 0.650 |
27
+ | Summarization | 0.767 |
28
+ | Instruction Following | 0.758 |
29
+
30
+ <div align="center">
31
+ <img width="70%" src="figures/fig3.png">
32
+ </div>
33
+
34
+ ## Training Details
35
+
36
+ The model was fine-tuned with an optimized hyperparameter configuration discovered through a systematic sweep. See the associated config.json for full details.
37
+
38
+ ## Usage
39
+
40
+ ```python
41
+ from transformers import AutoModelForSequenceClassification, AutoTokenizer
42
+
43
+ model = AutoModelForSequenceClassification.from_pretrained("FineTunedBest-TestRepo")
44
+ tokenizer = AutoTokenizer.from_pretrained("FineTunedBest-TestRepo")
45
+ ```
46
+
47
+ ## License
48
+
49
+ This model is released under the Apache 2.0 license.
config.json ADDED
@@ -0,0 +1,8 @@
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "model_type": "roberta",
3
+ "architectures": ["RobertaForSequenceClassification"],
4
+ "learning_rate": 5e-05,
5
+ "num_train_epochs": 5,
6
+ "batch_size": 8,
7
+ "seed": 21
8
+ }
figures/fig1.png ADDED
figures/fig2.png ADDED
figures/fig3.png ADDED
pytorch_model.bin ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:63fdef5f6fdfddc1e16513b3faab0bf7f823c1fb0d762c03a7e6f2fcfe88be63
3
+ size 28