Kuei-Chun Kao
Johnson0213
AI & ML interests
MLLM agent/ Reward model
Recent Activity
upvoted a paper about 15 hours ago
Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards authored a paper 11 days ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement upvoted a paper 11 days ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement