Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Runpeng Dai
Leo-Dai
6
37
2
Follow
TongZheng1999's profile picture
1 follower
·
3 following
AI & ML interests
None yet
Recent Activity
authored
a paper
11 days ago
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning
authored
a paper
11 days ago
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
upvoted
a
paper
11 days ago
Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation
View all activity
Organizations
Leo-Dai
's models
17
Sort: Recently updated
Leo-Dai/PPO_BL_250_critic
4B
•
Updated
Aug 15, 2025
•
2
Leo-Dai/PPO_BL_200_critic
Updated
Aug 15, 2025
•
2
Leo-Dai/PPO_BL_300_actor
Updated
Aug 15, 2025
•
4
Leo-Dai/PPO_BL_250_actor
Updated
Aug 15, 2025
•
2
Leo-Dai/PPO_BL_300_critic
Updated
Aug 15, 2025
Leo-Dai/GRPO_BL_40
4B
•
Updated
Aug 15, 2025
•
1
Leo-Dai/GRPO_BL_30
4B
•
Updated
Aug 15, 2025
•
4
Leo-Dai/GRPO_BL_20
4B
•
Updated
Aug 15, 2025
•
3
Leo-Dai/GRPO_BL_400
4B
•
Updated
Aug 15, 2025
•
2
Leo-Dai/GRPO_BL_10
4B
•
Updated
Aug 15, 2025
•
5
Leo-Dai/GRPO_BL_350
4B
•
Updated
Aug 15, 2025
•
1
Leo-Dai/GRPO_BL_200
4B
•
Updated
Aug 13, 2025
•
2
Leo-Dai/GRPO_BL_150
4B
•
Updated
Aug 13, 2025
•
2
Leo-Dai/GRPO_BL_100
4B
•
Updated
Aug 13, 2025
•
3
Leo-Dai/GRPO_BL_300
4B
•
Updated
Aug 13, 2025
•
3
Leo-Dai/GRPO_BL_250
4B
•
Updated
Aug 13, 2025
•
2
Leo-Dai/GRPO_BL_50
4B
•
Updated
Aug 13, 2025
•
2