Upload tokenizer

Files changed (5) hide show

source.spm ADDED Viewed

Binary file (805 kB). View file

special_tokens_map.json ADDED Viewed

+{
+  "eos_token": "</s>",
+  "pad_token": "<pad>",
+  "unk_token": "<unk>"
+}

target.spm ADDED Viewed

Binary file (807 kB). View file

tokenizer_config.json ADDED Viewed

+{
+  "eos_token": "</s>",
+  "model_max_length": 512,
+  "name_or_path": "/content/drive/MyDrive/translate_office_title/models/checkpoint-7000",
+  "pad_token": "<pad>",
+  "separate_vocabs": false,
+  "source_lang": "zho",
+  "sp_model_kwargs": {},
+  "special_tokens_map_file": null,
+  "target_lang": "eng",
+  "tokenizer_class": "MarianTokenizer",
+  "unk_token": "<unk>"
+}

vocab.json ADDED Viewed

The diff for this file is too large to render. See raw diff