Major Marc
ohhimarc
AI & ML interests
None yet
Recent Activity
updated a collection 24 days ago
ASR updated a collection about 1 month ago
OCR updated a collection about 1 month ago
OCROrganizations
None yet
Inference APIs
translation
LLMs
-
DiscoResearch/DiscoLM_German_7b_v1
Text Generation • 7B • Updated • 263 • • 67 -
DiscoResearch/Llama3-DiscoLeo-Instruct-8B-v0.1-4bit-awq
Text Generation • 8B • Updated • 36 -
google/gemma-2b-it
Text Generation • 3B • Updated • 24.6k • • 963 -
google/gemma-2-9b-it
Text Generation • 9B • Updated • 1.13M • • 1.04k
OCR
-
openbmb/MiniCPM-o-2_6
Any-to-Any • 9B • Updated • 330k • 1.3k -
microsoft/trocr-large-printed
Image-to-Text • 0.6B • Updated • 58.6k • 182 -
NAMAA-Space/Qari-OCR-0.2.2.1-VL-2B-Instruct
Image-Text-to-Text • Updated • 1.25k • 24 -
oddadmix/Qari-OCR-0.2.2.1-VL-2B-Instruct-merged
Image-Text-to-Text • 2B • Updated • 443 • 1
embeddings
-
sentence-transformers/paraphrase-multilingual-mpnet-base-v2
Sentence Similarity • 0.3B • Updated • 9.82M • • 525 -
sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2
Sentence Similarity • 0.1B • Updated • 46.3M • • 1.42k -
jinaai/jina-embeddings-v4
Visual Document Retrieval • 4B • Updated • 138k • 538 -
BAAI/bge-m3
Sentence Similarity • Updated • 35.9M • • 3.73k
cv
sentiment_analysis
-
citizenlab/distilbert-base-multilingual-cased-toxicity
Text Classification • Updated • 2.87k • • 23 -
oliverguhr/german-sentiment-bert
Text Classification • 0.1B • Updated • 276k • • 73 -
textdetox/xlmr-large-toxicity-classifier
Text Classification • 0.3B • Updated • 4.86k • • 17 -
gokceuludogan/convbert-base-turkish-mc4-toxicity-uncased
Text Classification • Updated • 117 • 3
forecasting
transcription
summarization
coding
ASR
-
openai/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4.32M • • 6.46k -
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.4M • • 3.4k -
nvidia/canary-1b
Automatic Speech Recognition • Updated • 2.78k • 459 -
badrex/mms-300m-arabic-dialect-identifier
Audio Classification • 0.3B • Updated • 955 • 9
NER
-
mdarhri00/named-entity-recognition
Token Classification • Updated • 90 • 55 -
eventdata-utd/conflibert-named-entity-recognition
Token Classification • 0.1B • Updated • 64 • 12 -
tomaarsen/span-marker-xlm-roberta-base-multinerd
Token Classification • 0.3B • Updated • 75 • 36 -
NAMAA-Space/gliner_arabic-v2.1
Token Classification • Updated • 210 • 21
text correction
sentiment_analysis
-
citizenlab/distilbert-base-multilingual-cased-toxicity
Text Classification • Updated • 2.87k • • 23 -
oliverguhr/german-sentiment-bert
Text Classification • 0.1B • Updated • 276k • • 73 -
textdetox/xlmr-large-toxicity-classifier
Text Classification • 0.3B • Updated • 4.86k • • 17 -
gokceuludogan/convbert-base-turkish-mc4-toxicity-uncased
Text Classification • Updated • 117 • 3
Inference APIs
forecasting
translation
transcription
LLMs
-
DiscoResearch/DiscoLM_German_7b_v1
Text Generation • 7B • Updated • 263 • • 67 -
DiscoResearch/Llama3-DiscoLeo-Instruct-8B-v0.1-4bit-awq
Text Generation • 8B • Updated • 36 -
google/gemma-2b-it
Text Generation • 3B • Updated • 24.6k • • 963 -
google/gemma-2-9b-it
Text Generation • 9B • Updated • 1.13M • • 1.04k
summarization
OCR
-
openbmb/MiniCPM-o-2_6
Any-to-Any • 9B • Updated • 330k • 1.3k -
microsoft/trocr-large-printed
Image-to-Text • 0.6B • Updated • 58.6k • 182 -
NAMAA-Space/Qari-OCR-0.2.2.1-VL-2B-Instruct
Image-Text-to-Text • Updated • 1.25k • 24 -
oddadmix/Qari-OCR-0.2.2.1-VL-2B-Instruct-merged
Image-Text-to-Text • 2B • Updated • 443 • 1
coding
embeddings
-
sentence-transformers/paraphrase-multilingual-mpnet-base-v2
Sentence Similarity • 0.3B • Updated • 9.82M • • 525 -
sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2
Sentence Similarity • 0.1B • Updated • 46.3M • • 1.42k -
jinaai/jina-embeddings-v4
Visual Document Retrieval • 4B • Updated • 138k • 538 -
BAAI/bge-m3
Sentence Similarity • Updated • 35.9M • • 3.73k
ASR
-
openai/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 4.32M • • 6.46k -
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.4M • • 3.4k -
nvidia/canary-1b
Automatic Speech Recognition • Updated • 2.78k • 459 -
badrex/mms-300m-arabic-dialect-identifier
Audio Classification • 0.3B • Updated • 955 • 9
cv
NER
-
mdarhri00/named-entity-recognition
Token Classification • Updated • 90 • 55 -
eventdata-utd/conflibert-named-entity-recognition
Token Classification • 0.1B • Updated • 64 • 12 -
tomaarsen/span-marker-xlm-roberta-base-multinerd
Token Classification • 0.3B • Updated • 75 • 36 -
NAMAA-Space/gliner_arabic-v2.1
Token Classification • Updated • 210 • 21