A Molecular Multimodal Foundation Model Associating Molecule Graphs with Natural Language Paper • 2209.05481 • Published Sep 12, 2022
CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning Paper • 2606.09393 • Published Jun 8
AdaGRPO: A Capability-Aware Adaptive Enhancement for Flow-based GRPO Paper • 2606.06828 • Published Jun 5
HPSD: Hybrid-Policy Self-Distillation for Text-Image-to-Video Diffusion Models Paper • 2608.13205 • Published 26 days ago • 3
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 5 days ago • 24
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 5 days ago • 24
JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence Paper • 2606.14777 • Published Jun 10 • 217
WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation Paper • 2605.10912 • Published May 11 • 48
HiFlow: Training-free High-Resolution Image Generation with Flow-Aligned Guidance Paper • 2504.06232 • Published Apr 8, 2025 • 13
Pref-GRPO: Pairwise Preference Reward-based GRPO for Stable Text-to-Image Reinforcement Learning Paper • 2508.20751 • Published Aug 28, 2025 • 91
UniGenBench++: A Unified Semantic Evaluation Benchmark for Text-to-Image Generation Paper • 2510.18701 • Published Oct 21, 2025 • 68
EndoCoT: Scaling Endogenous Chain-of-Thought Reasoning in Diffusion Models Paper • 2603.12252 • Published Mar 12 • 12
From Sparse to Dense: Multi-View GRPO for Flow Models via Augmented Condition Space Paper • 2603.12648 • Published Mar 13 • 14
From Sparse to Dense: Multi-View GRPO for Flow Models via Augmented Condition Space Paper • 2603.12648 • Published Mar 13 • 14
Running Agents 7 UniGenBench Leaderboard (English) 🏅 7 UniGenBench: a unified T2I generation benchmark.