Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Peng Shangpin's picture

Peng Shangpin

psp-dada
tencent
23 20 25
tangqh's profile picture FUJISyu0515's profile picture Tonic's profile picture
ยท
https://github.com/pspdada
  • pspdada
  • shangpin-peng-42a53b347

AI & ML interests

Multimodal Large Language Models, Preference Optimization Algorithm, Reinforcement Learning

Recent Activity

authored a paper 2 days ago
$T^5$: Twin-Critic Training for Token-Level Thoughts in Reinforcement Mid-Training
upvoted a paper 2 days ago
T^5: Twin-Critic Training for Token-Level Thoughts in Reinforcement Mid-Training
updated a dataset 4 days ago
psp-dada/TableVerse-5K
View all activity

Organizations

Tencent's profile picture
psp-dada 's papers 15
arxiv:2609.32791
arxiv:2608.12781
arxiv:2607.14548
arxiv:2607.04884
arxiv:2606.29905
arxiv:2606.23049
arxiv:2606.14832
arxiv:2606.01348
arxiv:2605.29486
arxiv:2605.11960
arxiv:2605.07630
arxiv:2605.03677
arxiv:2506.10054
arxiv:2511.19575
arxiv:2507.12455
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs