Audio Dataset MLCommons/peoples_speech_v1.0 Updated Aug 25, 2024 • 45.3k • 8 amphion/Emilia-Dataset Viewer • Updated Feb 28, 2025 • 54.8M • 93.3k • 472 simon3000/genshin-voice Viewer • Updated 8 days ago • 631k • 7.28k • 246 facebook/multilingual_librispeech Viewer • Updated Aug 12, 2024 • 1.49M • 19.8k • 184
Omni model collection of Omni modal model inclusionAI/Ming-flash-omni-2.0 Any-to-Any • 104B • Updated Feb 12 • 2.68k • 267 Qwen/Qwen3-Omni-30B-A3B-Instruct Any-to-Any • 35B • Updated Sep 22, 2025 • 1.5M • 963 naver-hyperclovax/HyperCLOVAX-SEED-Omni-8B Text Generation • 11B • Updated Jan 6 • 313 • 192 meituan-longcat/LongCat-Flash-Omni Any-to-Any • 561B • Updated Nov 11, 2025 • 79 • 115
Audio Dataset MLCommons/peoples_speech_v1.0 Updated Aug 25, 2024 • 45.3k • 8 amphion/Emilia-Dataset Viewer • Updated Feb 28, 2025 • 54.8M • 93.3k • 472 simon3000/genshin-voice Viewer • Updated 8 days ago • 631k • 7.28k • 246 facebook/multilingual_librispeech Viewer • Updated Aug 12, 2024 • 1.49M • 19.8k • 184
Omni model collection of Omni modal model inclusionAI/Ming-flash-omni-2.0 Any-to-Any • 104B • Updated Feb 12 • 2.68k • 267 Qwen/Qwen3-Omni-30B-A3B-Instruct Any-to-Any • 35B • Updated Sep 22, 2025 • 1.5M • 963 naver-hyperclovax/HyperCLOVAX-SEED-Omni-8B Text Generation • 11B • Updated Jan 6 • 313 • 192 meituan-longcat/LongCat-Flash-Omni Any-to-Any • 561B • Updated Nov 11, 2025 • 79 • 115