Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Xiangxin Zhou
zhouxiangxin
3
21
4
Follow
JohnRoger's profile picture
Gargaz's profile picture
Datawitch-Programmer's profile picture
6 followers
·
13 following
https://zhouxiangxin1998.github.io/
AI & ML interests
None yet
Recent Activity
authored
a paper
about 1 month ago
Rethinking the Divergence Regularization in LLM RL
authored
a paper
about 1 month ago
Flow-DPPO: Divergence Proximal Policy Optimization for Flow Matching Models
authored
a paper
about 1 month ago
Beyond Uniform Token-Level Trust Region in LLM Reinforcement Learning
View all activity
Organizations
zhouxiangxin
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
commented
2 papers
about 1 month ago
Rethinking the Divergence Regularization in LLM RL
Paper
•
2606.09821
•
Published
Jun 8
•
34
•
4
Rethinking the Divergence Regularization in LLM RL
Paper
•
2606.09821
•
Published
Jun 8
•
34
•
4