Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Nguyễn Minh Phúc
DatPySci
7
1
Follow
Oztobuzz's profile picture
dark-pen's profile picture
2 followers
·
2 following
AI & ML interests
Reinforcement learning, NLP
Recent Activity
upvoted
an
article
3 days ago
Navigating the RLHF Landscape: From Policy Gradients to PPO, GAE, and DPO for LLM Alignment
updated
a model
6 days ago
DatPySci/Checkpoints
published
a model
10 days ago
DatPySci/Checkpoints
View all activity
Organizations
DatPySci
's models
6
Sort: Recently updated
DatPySci/Checkpoints
Updated
5 days ago
DatPySci/check
Updated
17 days ago
•
1
DatPySci/AR-Diff
Updated
21 days ago
DatPySci/LazyNTK
Updated
Apr 27
DatPySci/RLVR-CoTs
Updated
Feb 26
DatPySci/RLDI
2B
•
Updated
Dec 18, 2025
•
16