Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
黄炜锴
tsrigo
15
4
Follow
0 followers
·
2 following
tsrigo
AI & ML interests
Trustworthy AI
Recent Activity
upvoted
a
paper
1 day ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning
upvoted
a
paper
2 days ago
Progressive Agent Skill Generation via Reinforcement Learning
upvoted
a
paper
9 days ago
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search
View all activity
Organizations
models
4
Sort: Recently updated
tsrigo/coconut
Updated
Feb 15, 2025
•
1
tsrigo/unsloth_model
Text Generation
•
3B
•
Updated
Feb 13, 2025
•
8
tsrigo/Qwen2.5-1.5B-Instruct-DPO-bad-boy
2B
•
Updated
Jan 20, 2025
•
5
tsrigo/Qwen2.5-0.5B-Instruct-DPO-bad-boy
Updated
Jan 20, 2025
datasets
1
tsrigo/btfChinese-DPO-small
Viewer
•
Updated
Jan 20, 2025
•
5k
•
47