arxiv:2606.13679
Dian Zheng PRO
zhengli1013
AI & ML interests
generative model
Recent Activity
upvoted a paper about 4 hours ago
TACD: Distilling Efficient Text-to-Motion Models via Terminal Amplification Control upvoted a paper about 1 month ago
Puffin-World: Scaling a Unified Multimodal Model with Native 3D World States upvoted a paper about 1 month ago
Video-IFBench: Evaluating Instruction Following of Multimodal LLMs in Video Understanding Scenarios