LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 3 days ago • 209
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 6 days ago • 97
FunAudioLLM/Fun-ASR-Nano-2512-hf Automatic Speech Recognition • 0.8B • Updated 3 days ago • 21.6k • 15
KVAE: Family of Tokenizers for Multimodal Generative Models Paper • 2608.05798 • Published about 1 month ago • 29