๐ In a Training Loop
Juntong Fang
hellexf
ยท
AI & ML interests
None yet
Recent Activity
upvoted a paper 20 days ago
DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation upvoted a paper 30 days ago
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning upvoted a paper about 1 month ago
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text SpaceOrganizations
None yet