badokorach
/

distilbert-base-cased-distilled-agric-trans-240224

@@ -1,6 +1,6 @@
 ---
 license: apache-2.0
-base_model: badokorach/distilbert-base-cased-distilled-agric-060124_1
 tags:
 - generated_from_keras_callback
 model-index:
@@ -13,11 +13,11 @@ probably proofread and complete it, then remove this comment. -->
 # badokorach/distilbert-base-cased-distilled-agric-trans-240224
-This model is a fine-tuned version of [badokorach/distilbert-base-cased-distilled-agric-060124_1](https://huggingface.co/badokorach/distilbert-base-cased-distilled-agric-060124_1) on an unknown dataset.
 It achieves the following results on the evaluation set:
-- Train Loss: 0.1281
 - Validation Loss: 0.0
-- Epoch: 14
 ## Model description
@@ -36,28 +36,14 @@ More information needed
 ### Training hyperparameters
 The following hyperparameters were used during training:
-- optimizer: {'name': 'AdamWeightDecay', 'learning_rate': {'module': 'keras.optimizers.schedules', 'class_name': 'PolynomialDecay', 'config': {'initial_learning_rate': 1e-05, 'decay_steps': 690, 'end_learning_rate': 0.0, 'power': 1.0, 'cycle': False, 'name': None}, 'registered_name': None}, 'decay': 0.0, 'beta_1': 0.9, 'beta_2': 0.999, 'epsilon': 1e-08, 'amsgrad': False, 'weight_decay_rate': 0.02}
 - training_precision: mixed_float16
 ### Training results
 | Train Loss | Validation Loss | Epoch |
 |:----------:|:---------------:|:-----:|
-| 0.5050     | 0.0             | 0     |
-| 0.4219     | 0.0             | 1     |
-| 0.3805     | 0.0             | 2     |
-| 0.3095     | 0.0             | 3     |
-| 0.2684     | 0.0             | 4     |
-| 0.2508     | 0.0             | 5     |
-| 0.2261     | 0.0             | 6     |
-| 0.2027     | 0.0             | 7     |
-| 0.1752     | 0.0             | 8     |
-| 0.1588     | 0.0             | 9     |
-| 0.1535     | 0.0             | 10    |
-| 0.1418     | 0.0             | 11    |
-| 0.1394     | 0.0             | 12    |
-| 0.1347     | 0.0             | 13    |
-| 0.1281     | 0.0             | 14    |
 ### Framework versions

 ---
 license: apache-2.0
+base_model: distilbert-base-cased-distilled-squad
 tags:
 - generated_from_keras_callback
 model-index:
 # badokorach/distilbert-base-cased-distilled-agric-trans-240224
+This model is a fine-tuned version of [distilbert-base-cased-distilled-squad](https://huggingface.co/distilbert-base-cased-distilled-squad) on an unknown dataset.
 It achieves the following results on the evaluation set:
+- Train Loss: 2.8831
 - Validation Loss: 0.0
+- Epoch: 0
 ## Model description
 ### Training hyperparameters
 The following hyperparameters were used during training:
+- optimizer: {'inner_optimizer': {'module': 'transformers.optimization_tf', 'class_name': 'AdamWeightDecay', 'config': {'name': 'AdamWeightDecay', 'learning_rate': {'module': 'keras.optimizers.schedules', 'class_name': 'PolynomialDecay', 'config': {'initial_learning_rate': 1e-05, 'decay_steps': 690, 'end_learning_rate': 0.0, 'power': 1.0, 'cycle': False, 'name': None}, 'registered_name': None}, 'decay': 0.0, 'beta_1': 0.8999999761581421, 'beta_2': 0.9990000128746033, 'epsilon': 1e-08, 'amsgrad': False, 'weight_decay_rate': 0.02}, 'registered_name': 'AdamWeightDecay'}, 'dynamic': True, 'initial_scale': 32768.0, 'dynamic_growth_steps': 2000}
 - training_precision: mixed_float16
 ### Training results
 | Train Loss | Validation Loss | Epoch |
 |:----------:|:---------------:|:-----:|
+| 2.8831     | 0.0             | 0     |
 ### Framework versions

config.json CHANGED Viewed

@@ -1,5 +1,5 @@
 {
-  "_name_or_path": "badokorach/distilbert-base-cased-distilled-agric-060124_1",
   "activation": "gelu",
   "architectures": [
     "DistilBertForQuestionAnswering"

 {
+  "_name_or_path": "distilbert-base-cased-distilled-squad",
   "activation": "gelu",
   "architectures": [
     "DistilBertForQuestionAnswering"

special_tokens_map.json CHANGED Viewed

@@ -1,37 +1,7 @@
 {
-  "cls_token": {
-    "content": "[CLS]",
-    "lstrip": false,
-    "normalized": false,
-    "rstrip": false,
-    "single_word": false
-  },
-  "mask_token": {
-    "content": "[MASK]",
-    "lstrip": false,
-    "normalized": false,
-    "rstrip": false,
-    "single_word": false
-  },
-  "pad_token": {
-    "content": "[PAD]",
-    "lstrip": false,
-    "normalized": false,
-    "rstrip": false,
-    "single_word": false
-  },
-  "sep_token": {
-    "content": "[SEP]",
-    "lstrip": false,
-    "normalized": false,
-    "rstrip": false,
-    "single_word": false
-  },
-  "unk_token": {
-    "content": "[UNK]",
-    "lstrip": false,
-    "normalized": false,
-    "rstrip": false,
-    "single_word": false
-  }
 }

 {
+  "cls_token": "[CLS]",
+  "mask_token": "[MASK]",
+  "pad_token": "[PAD]",
+  "sep_token": "[SEP]",
+  "unk_token": "[UNK]"
 }

tf_model.h5 CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:9c4783025a57eb521772cb2497dacd4db62fdd834c606bacdca2c2617f7ac540
 size 260895720

 version https://git-lfs.github.com/spec/v1
+oid sha256:2718d2585fc9daf08c5831aad44a390c4f51e2c12f0ab6d61f3ea7c73e79f3c9
 size 260895720