Kuei-Chun Kao
Johnson0213
AI & ML interests
MLLM agent/ Reward model
Recent Activity
upvoted a paper about 23 hours ago
Post-Training Frontier Text-to-Image Models by Composing Preference and Rubric Rewards authored a paper 12 days ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement upvoted a paper 12 days ago
ReCAST: Reward Credit Assignment across Timesteps for Online Diffusion Reinforcement