A first approach for general audio generation with high-dimensional LLM + Diffusion.
AI & ML interests
None defined yet.
Recent Activity
Organization Card
models 26
mispeech/midashenglm-gen
Text-to-Audio • 3B • Updated • 185 • 45
mispeech/Dasheng-AudioGen
Text-to-Audio • 2B • Updated • 590 • 18
mispeech/Dasheng-AudioGen-Multilingual
Text-to-Audio • 2B • Updated • 53 • 6
mispeech/dasheng-denoiser
Audio-to-Audio • 0.1B • Updated • 98 • 15
mispeech/dashengtokenizer
Audio-to-Audio • 0.8B • Updated • 833 • 15
mispeech/midashenglm-0.6b-gguf
Audio-Text-to-Text • 0.6B • Updated • 585 • 1
mispeech/midashenglm-7b-1021-gguf
Audio-Text-to-Text • 8B • Updated • 900 • 3
mispeech/midashenglm-0.6b-fp32
Audio-Text-to-Text • 0.7B • Updated • 543 • 4
mispeech/ced-base
Audio Classification • 85.7M • Updated • 20.3k • 16
mispeech/ced-tiny
Audio Classification • 5.5M • Updated • 2.8k • 5