arxiv:2609.13425
Kuei-Chun Kao
Johnson0213
AI & ML interests
MLLM agent/ Reward model
Recent Activity
authored a paper about 9 hours ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement upvoted a paper about 15 hours ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement upvoted a paper 8 days ago
QG-CoC: Question-Guided Chain-of-Captions for Large Multimodal Models