Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
li sheng
bambisheng
1
21
4
Follow
John6666's profile picture
Farhad9849's profile picture
XingtaiHF's profile picture
5 followers
·
14 following
https://github.com/BambiSheng
AI & ML interests
None yet
Recent Activity
upvoted
a
paper
about 10 hours ago
EasyPPO: Stabilizing the Critic Is Key
upvoted
a
paper
7 days ago
Improving Test-Time Scaling with Adaptive Looped Transformers
upvoted
a
paper
25 days ago
T1: Terminal Agent Reinforcement Learning for Long-Horizon Tasks
View all activity
Organizations
bambisheng
's models
3
Sort: Recently updated
bambisheng/UltraIF-8B-DPO
Text Generation
•
8B
•
Updated
Apr 3, 2025
•
16
•
3
bambisheng/UltraIF-8B-UltraComposer
Text Generation
•
8B
•
Updated
Apr 3, 2025
•
25
•
1
bambisheng/UltraIF-8B-SFT
Text Generation
•
8B
•
Updated
Apr 3, 2025
•
16
•
2