ckpt and expert models(dev) for flow-opd.
Zhen Fang
CostaliyA
AI & ML interests
None yet
Recent Activity
upvoted a paper about 2 hours ago
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation upvoted a paper about 19 hours ago
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement LearningOrganizations
None yet