A first approach for general audio generation with high-dimensional LLM + Diffusion.
AI & ML interests
None defined yet.
Recent Activity
Organization Card
models 26
mispeech/midashenglm-gen
Text-to-Audio • 3B • Updated • 995 • 41
mispeech/Dasheng-AudioGen
Text-to-Audio • 2B • Updated • 762 • 15
mispeech/Dasheng-AudioGen-Multilingual
Text-to-Audio • 2B • Updated • 59 • 6
mispeech/dasheng-denoiser
Audio-to-Audio • 0.1B • Updated • 76 • 15
mispeech/dashengtokenizer
Audio-to-Audio • 0.8B • Updated • 5.45k • 12
mispeech/midashenglm-0.6b-gguf
Audio-Text-to-Text • 0.6B • Updated • 181 • 1
mispeech/midashenglm-7b-1021-gguf
Audio-Text-to-Text • 8B • Updated • 224 • 3
mispeech/midashenglm-0.6b-fp32
Audio-Text-to-Text • 0.7B • Updated • 239 • 4
mispeech/ced-base
Audio Classification • 85.7M • Updated • 12.2k • 15
mispeech/ced-tiny
Audio Classification • 5.5M • Updated • 3.4k • 4