Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Peng Shangpin
psp-dada
23
20
25
Follow
tangqh's profile picture
Tonic's profile picture
FUJISyu0515's profile picture
3 followers
·
4 following
https://github.com/pspdada
pspdada
shangpin-peng-42a53b347
AI & ML interests
Multimodal Large Language Models, Preference Optimization Algorithm, Reinforcement Learning
Recent Activity
authored
a paper
1 day ago
$T^5$: Twin-Critic Training for Token-Level Thoughts in Reinforcement Mid-Training
upvoted
a
paper
1 day ago
T^5: Twin-Critic Training for Token-Level Thoughts in Reinforcement Mid-Training
updated
a dataset
3 days ago
psp-dada/TableVerse-5K
View all activity
Organizations
psp-dada
's datasets
4
Sort: Recently updated
psp-dada/TableVerse-5K
Viewer
•
Updated
3 days ago
•
4.84k
•
541
psp-dada/SENTINEL
Updated
Jul 2
•
257
•
2
psp-dada/ChartArena
Viewer
•
Updated
Jun 14
•
2.39k
•
347
psp-dada/Uni-DPO
Preview
•
Updated
Feb 22
•
84
•
1