ZhengHao
ZhengHao-L
AI & ML interests
None yet
Recent Activity
upvoted a paper 27 days ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement upvoted a paper about 2 months ago
SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring upvoted a paper about 2 months ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-ImprovementOrganizations
None yet