DNeff/laya-bucket / tokenizer /tokenizer_config.json
DNeff's picture
download
raw
308 Bytes
{
"clean_up_tokenization_spaces": true,
"cls_token": "[CLS]",
"mask_token": "[MASK]",
"model_input_names": [
"input_ids",
"attention_mask"
],
"model_max_length": 8192,
"pad_token": "[PAD]",
"sep_token": "[SEP]",
"tokenizer_class": "PreTrainedTokenizerFast",
"unk_token": "[UNK]"
}

Xet Storage Details

Size:
308 Bytes
·
Xet hash:
972b76662a6f91e8a559a7de5163e7281947cb1c05f5680ab3f66c3c50ecdc2b

Xet efficiently stores files, intelligently splitting them into unique chunks and accelerating uploads and downloads. More info.