Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
gbyuvd
/
FastChemTokenizer
Like
0
Feature Extraction
qwen3
chemistry
tokenizer
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
FastChemTokenizer
9.96 MB
Ctrl+K
Ctrl+K
1 contributor
History:
41 commits
gbyuvd
Update FastChemTokenizerHF2.py
8cc4c16
verified
12 months ago
benchmark
Upload latent visualization notebook
about 1 year ago
bigsmiles-proto
Upload BigSMILES vocab
about 1 year ago
latent_space_plots
Upload benchmark script and set
about 1 year ago
selftok_core
Update to include SELFIES Tokenizer & Vocabs
about 1 year ago
selftok_wtails
Update to include SELFIES Tokenizer & Vocabs
about 1 year ago
smitok
First commit
about 1 year ago
smitok_core
Upload HF wrapper and smitok_core without tails
about 1 year ago
.gitattributes
Safe
1.71 kB
Upload benchmark script and set
about 1 year ago
CHANGELOG
Safe
193 Bytes
Tensor handling fix
about 1 year ago
FastChemTokenizer.py
Safe
23.1 kB
Tensor handling fix
about 1 year ago
FastChemTokenizerHF.py
Safe
24 kB
Proper full HF Compat
12 months ago
FastChemTokenizerHF2.py
25.5 kB
Update FastChemTokenizerHF2.py
12 months ago
README.md
Safe
12.2 kB
Update README.md
about 1 year ago
config.json
Safe
896 Bytes
Update config.json
about 1 year ago
requirements.txt
Safe
120 Bytes
Upload requirements.txt
about 1 year ago