Jingzhi Wang
jzwang666
ยท
AI & ML interests
LLM pretrain & finetuning, develop high effiency attention architecture, agentic RL system
Recent Activity
authored a paper about 4 hours ago
MemSFT: Mitigating Alignment Tax with an External Parametric Memory authored a paper about 4 hours ago
JTok: On Token Embedding as another Axis of Scaling Law via Joint Token Self-modulation authored a paper about 4 hours ago
ssToken: Self-modulated and Semantic-aware Token Selection for LLM
Fine-tuning