Verified Story Studio creative-audio models and runtimes for Vox Jot.
Kimani James
IrieDinamik
·
AI & ML interests
None yet
Recent Activity
updated a model 13 days ago
IrieDinamik/vox-jot-models updated a model 14 days ago
IrieDinamik/vox-jot-releases updated a model 27 days ago
IrieDinamik/vox-jot-speech-analysis-runtimeOrganizations
None yet
Vox Jot – Speech Analysis Runtime
Public managed runtime for Vox Jot file ASR, diarization, Polyvoice routing, and emotion analysis.
Vox Jot – Speaker Isolation Verified
Curated speaker diarization and isolation models verified for Vox Jot file transcription.
-
pyannote/speaker-diarization-community-1
Automatic Speech Recognition • Updated • 5.35M • 1.01k -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 8.87M • 3.06k -
BUT-FIT/diarizen-wavlm-large-s80-md-v2
Voice Activity Detection • Updated • 2.14k • 19 -
nvidia/diar_sortformer_4spk-v1
Automatic Speech Recognition • 0.1B • Updated • 19.8k • 151
Vox Jot – OCR Verified
Curated on-device OCR models verified for Vox Jot. Image-to-text and document scanning models optimized for local inference on-device.
Vox Jot – TTS Verified
Curated on-device TTS models verified for Vox Jot speech synthesis. Ranked: Fastest → Balanced → Best Quality → Voice Cloning. MIT/Apache licensed.
Vox Jot - TTS Candidates
Candidate on-device TTS models under evaluation for Vox Jot. Models here are not ranked or verified until full Vox Jot benchmark suites pass.
ML Models
Machine Learning Models
Vox Jot – File ASR Verified
Curated file-transcription ASR models verified for Vox Jot. File/audio engines, not live dictation hot path.
-
ibm-granite/granite-speech-4.1-2b
Automatic Speech Recognition • 2B • Updated • 338k • 158 -
CohereLabs/cohere-transcribe-03-2026
Automatic Speech Recognition • 2B • Updated • 868k • • 1.08k -
Systran/faster-whisper-large-v3
Automatic Speech Recognition • Updated • 1.07M • 635 -
mlx-community/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 896 • 11
Vox Jot – LLM Verified
Curated on-device LLM/GGUF models verified for Vox Jot. Small, fast instruct models optimized for local inference on-device.
-
IrieDinamik/LiquidAI-LFM2.5-Audio-1.5B-GGUF
1B • Updated • 104 -
IrieDinamik/LiquidAI-LFM2-1.2B-Tool-GGUF
Text Generation • 1B • Updated • 49 -
bartowski/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 135k • 230 -
bartowski/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 134k • 173
Vox Jot – STT Verified
Curated CTranslate2 Whisper models verified for Vox Jot speech-to-text. Ranked: Fastest → Balanced → Best Quality → Experimental. MIT/Apache licensed,
-
Systran/faster-whisper-tiny
Automatic Speech Recognition • Updated • 1.33M • 24 -
Systran/faster-whisper-tiny.en
Automatic Speech Recognition • Updated • 1.22M • 10 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.45M • 32 -
Systran/faster-whisper-base.en
Automatic Speech Recognition • Updated • 218k • 7
Vox Jot - Creative Audio Verified
Verified Story Studio creative-audio models and runtimes for Vox Jot.
Vox Jot - TTS Candidates
Candidate on-device TTS models under evaluation for Vox Jot. Models here are not ranked or verified until full Vox Jot benchmark suites pass.
Vox Jot – Speech Analysis Runtime
Public managed runtime for Vox Jot file ASR, diarization, Polyvoice routing, and emotion analysis.
ML Models
Machine Learning Models
Vox Jot – Speaker Isolation Verified
Curated speaker diarization and isolation models verified for Vox Jot file transcription.
-
pyannote/speaker-diarization-community-1
Automatic Speech Recognition • Updated • 5.35M • 1.01k -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 8.87M • 3.06k -
BUT-FIT/diarizen-wavlm-large-s80-md-v2
Voice Activity Detection • Updated • 2.14k • 19 -
nvidia/diar_sortformer_4spk-v1
Automatic Speech Recognition • 0.1B • Updated • 19.8k • 151
Vox Jot – File ASR Verified
Curated file-transcription ASR models verified for Vox Jot. File/audio engines, not live dictation hot path.
-
ibm-granite/granite-speech-4.1-2b
Automatic Speech Recognition • 2B • Updated • 338k • 158 -
CohereLabs/cohere-transcribe-03-2026
Automatic Speech Recognition • 2B • Updated • 868k • • 1.08k -
Systran/faster-whisper-large-v3
Automatic Speech Recognition • Updated • 1.07M • 635 -
mlx-community/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 896 • 11
Vox Jot – OCR Verified
Curated on-device OCR models verified for Vox Jot. Image-to-text and document scanning models optimized for local inference on-device.
Vox Jot – LLM Verified
Curated on-device LLM/GGUF models verified for Vox Jot. Small, fast instruct models optimized for local inference on-device.
-
IrieDinamik/LiquidAI-LFM2.5-Audio-1.5B-GGUF
1B • Updated • 104 -
IrieDinamik/LiquidAI-LFM2-1.2B-Tool-GGUF
Text Generation • 1B • Updated • 49 -
bartowski/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 135k • 230 -
bartowski/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 134k • 173
Vox Jot – TTS Verified
Curated on-device TTS models verified for Vox Jot speech synthesis. Ranked: Fastest → Balanced → Best Quality → Voice Cloning. MIT/Apache licensed.
Vox Jot – STT Verified
Curated CTranslate2 Whisper models verified for Vox Jot speech-to-text. Ranked: Fastest → Balanced → Best Quality → Experimental. MIT/Apache licensed,
-
Systran/faster-whisper-tiny
Automatic Speech Recognition • Updated • 1.33M • 24 -
Systran/faster-whisper-tiny.en
Automatic Speech Recognition • Updated • 1.22M • 10 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.45M • 32 -
Systran/faster-whisper-base.en
Automatic Speech Recognition • Updated • 218k • 7