Upload tokenizer

Files changed (5) hide show

source.spm ADDED Viewed

Binary file (797 kB). View file

special_tokens_map.json ADDED Viewed

+{
+  "eos_token": "</s>",
+  "pad_token": "<pad>",
+  "unk_token": "<unk>"
+}

target.spm ADDED Viewed

Binary file (768 kB). View file

tokenizer_config.json ADDED Viewed

+{
+  "eos_token": "</s>",
+  "model_max_length": 512,
+  "name_or_path": "fine-tune/checkpoint-250000-finetuned-lb-to-en-deen-120k-flores-500/checkpoint-500",
+  "pad_token": "<pad>",
+  "separate_vocabs": false,
+  "source_lang": "de",
+  "sp_model_kwargs": {},
+  "special_tokens_map_file": null,
+  "target_lang": "en",
+  "tokenizer_class": "MarianTokenizer",
+  "unk_token": "<unk>"
+}

vocab.json ADDED Viewed

The diff for this file is too large to render. See raw diff