herbert-klej-cased-tokenizer-v1 / tokenizer_config.json
system's picture
system HF staff
Update tokenizer_config.json
b78be7b
raw
history blame
341 Bytes
{"do_lowercase_and_remove_accent": false, "cls_token": "<s>", "bos_token": "<s>", "additional_special_tokens": ["<special0>", "<special1>", "<special2>", "<special3>", "<special4>", "<special5>", "<special6>", "<special7>", "<special8>", "<special9>"], "pad_token": "<pad>", "sep_token": "</s>", "mask_token": "<mask>", "unk_token": "<unk>"}