Verified Story Studio creative-audio models and runtimes for Vox Jot.
Kimani James
IrieDinamik
·
AI & ML interests
None yet
Recent Activity
updated a model about 21 hours ago
IrieDinamik/vox-jot-models updated a model about 23 hours ago
IrieDinamik/vox-jot-releases updated a model 14 days ago
IrieDinamik/vox-jot-speech-analysis-runtimeOrganizations
None yet
Vox Jot – Speech Analysis Runtime
Public managed runtime for Vox Jot file ASR, diarization, Polyvoice routing, and emotion analysis.
Vox Jot – Speaker Isolation Verified
Curated speaker diarization and isolation models verified for Vox Jot file transcription.
-
pyannote/speaker-diarization-community-1
Automatic Speech Recognition • Updated • 4.91M • 896 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 8.62M • 2.93k -
BUT-FIT/diarizen-wavlm-large-s80-md-v2
Voice Activity Detection • Updated • 2.69k • 19 -
nvidia/diar_sortformer_4spk-v1
Automatic Speech Recognition • 0.1B • Updated • 11.7k • 150
Vox Jot – OCR Verified
Curated on-device OCR models verified for Vox Jot. Image-to-text and document scanning models optimized for local inference on-device.
Vox Jot – TTS Verified
Curated on-device TTS models verified for Vox Jot speech synthesis. Ranked: Fastest → Balanced → Best Quality → Voice Cloning. MIT/Apache licensed.
Vox Jot - TTS Candidates
Candidate on-device TTS models under evaluation for Vox Jot. Models here are not ranked or verified until full Vox Jot benchmark suites pass.
ML Models
Machine Learning Models
Vox Jot – File ASR Verified
Curated file-transcription ASR models verified for Vox Jot. File/audio engines, not live dictation hot path.
-
ibm-granite/granite-speech-4.1-2b
Automatic Speech Recognition • 2B • Updated • 467k • 153 -
CohereLabs/cohere-transcribe-03-2026
Automatic Speech Recognition • 2B • Updated • 1.03M • • 1.07k -
Systran/faster-whisper-large-v3
Automatic Speech Recognition • Updated • 1.19M • 628 -
mlx-community/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 935 • 11
Vox Jot – LLM Verified
Curated on-device LLM/GGUF models verified for Vox Jot. Small, fast instruct models optimized for local inference on-device.
-
IrieDinamik/LiquidAI-LFM2.5-Audio-1.5B-GGUF
1B • Updated • 186 -
IrieDinamik/LiquidAI-LFM2-1.2B-Tool-GGUF
Text Generation • 1B • Updated • 102 -
bartowski/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 157k • 229 -
bartowski/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 162k • 172
Vox Jot – STT Verified
Curated CTranslate2 Whisper models verified for Vox Jot speech-to-text. Ranked: Fastest → Balanced → Best Quality → Experimental. MIT/Apache licensed,
-
Systran/faster-whisper-tiny
Automatic Speech Recognition • Updated • 1.36M • 24 -
Systran/faster-whisper-tiny.en
Automatic Speech Recognition • Updated • 1.16M • 10 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.44M • 32 -
Systran/faster-whisper-base.en
Automatic Speech Recognition • Updated • 189k • 7
Vox Jot - Creative Audio Verified
Verified Story Studio creative-audio models and runtimes for Vox Jot.
Vox Jot - TTS Candidates
Candidate on-device TTS models under evaluation for Vox Jot. Models here are not ranked or verified until full Vox Jot benchmark suites pass.
Vox Jot – Speech Analysis Runtime
Public managed runtime for Vox Jot file ASR, diarization, Polyvoice routing, and emotion analysis.
ML Models
Machine Learning Models
Vox Jot – Speaker Isolation Verified
Curated speaker diarization and isolation models verified for Vox Jot file transcription.
-
pyannote/speaker-diarization-community-1
Automatic Speech Recognition • Updated • 4.91M • 896 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition • Updated • 8.62M • 2.93k -
BUT-FIT/diarizen-wavlm-large-s80-md-v2
Voice Activity Detection • Updated • 2.69k • 19 -
nvidia/diar_sortformer_4spk-v1
Automatic Speech Recognition • 0.1B • Updated • 11.7k • 150
Vox Jot – File ASR Verified
Curated file-transcription ASR models verified for Vox Jot. File/audio engines, not live dictation hot path.
-
ibm-granite/granite-speech-4.1-2b
Automatic Speech Recognition • 2B • Updated • 467k • 153 -
CohereLabs/cohere-transcribe-03-2026
Automatic Speech Recognition • 2B • Updated • 1.03M • • 1.07k -
Systran/faster-whisper-large-v3
Automatic Speech Recognition • Updated • 1.19M • 628 -
mlx-community/nemotron-3.5-asr-streaming-0.6b
Automatic Speech Recognition • 0.6B • Updated • 935 • 11
Vox Jot – OCR Verified
Curated on-device OCR models verified for Vox Jot. Image-to-text and document scanning models optimized for local inference on-device.
Vox Jot – LLM Verified
Curated on-device LLM/GGUF models verified for Vox Jot. Small, fast instruct models optimized for local inference on-device.
-
IrieDinamik/LiquidAI-LFM2.5-Audio-1.5B-GGUF
1B • Updated • 186 -
IrieDinamik/LiquidAI-LFM2-1.2B-Tool-GGUF
Text Generation • 1B • Updated • 102 -
bartowski/Llama-3.2-3B-Instruct-GGUF
Text Generation • 3B • Updated • 157k • 229 -
bartowski/Llama-3.2-1B-Instruct-GGUF
Text Generation • 1B • Updated • 162k • 172
Vox Jot – TTS Verified
Curated on-device TTS models verified for Vox Jot speech synthesis. Ranked: Fastest → Balanced → Best Quality → Voice Cloning. MIT/Apache licensed.
Vox Jot – STT Verified
Curated CTranslate2 Whisper models verified for Vox Jot speech-to-text. Ranked: Fastest → Balanced → Best Quality → Experimental. MIT/Apache licensed,
-
Systran/faster-whisper-tiny
Automatic Speech Recognition • Updated • 1.36M • 24 -
Systran/faster-whisper-tiny.en
Automatic Speech Recognition • Updated • 1.16M • 10 -
Systran/faster-whisper-base
Automatic Speech Recognition • Updated • 1.44M • 32 -
Systran/faster-whisper-base.en
Automatic Speech Recognition • Updated • 189k • 7