I am currently a Ph.D. student at Sun Yat-sen University, specializing in Humanoid Learning and Video Understanding.
My research interests focus on humanoid learning, action recognition, and temporal video understanding. I am interested in building robust embodied and video understanding systems and aim to make meaningful contributions through sustained research.
Before starting my Ph.D., I was an undergraduate student at South China University of Technology studying Network Engineering. During my undergraduate studies, I was actively involved in algorithm competitions and research internships, gaining experience in computer vision and machine learning. I have received several honors, including the President Scholarship at Sun Yat-sen University, the National Scholarship, Kaggle Silver medals, and various competition prizes.
Feel free to reach out for research collaborations or academic discussions!
Our paper “ChoreoPlan: Hybrid Phrase Planning and Execution-Grounded Selection for Music-to-Humanoid Dance” was accepted to ACM MM 2026.
Feb 20, 2026
Our paper “Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations” was accepted to CVPR 2026.
Jun 25, 2025
Our paper “Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning” was accepted to ICCV 2025.
Feb 26, 2025
Our paper “Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks” was accepted to CVPR 2025.
Sep 06, 2024
Updated personal website content and contact email to huangwj235@mail2.sysu.edu.cn.
@inproceedings{huang2026choreoplan,title={ChoreoPlan: Hybrid Phrase Planning and Execution-Grounded Selection for Music-to-Humanoid Dance},author={Huang, Wei-Jin and Fan, Jianhong and Huang, Hao and Xia, Zhi-Wei and Deng, Jun-Yi and Li, Yuan-Ming and Lin, Kun-Yu and Zheng, Wei-Shi},booktitle={Proceedings of the ACM International Conference on Multimedia},year={2026},}
Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations
CVPR
Wei-Jin Huang , Yue-Yi Zhang , Yi-Lin Wei , Zhi-Wei Xia , Juantao Tan , Yuan-Ming Li , Zhilin Zhao , and Wei-Shi Zheng
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Jun 2026
@inproceedings{huang2026learning,title={Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations},author={Huang, Wei-Jin and Zhang, Yue-Yi and Wei, Yi-Lin and Xia, Zhi-Wei and Tan, Juantao and Li, Yuan-Ming and Zhao, Zhilin and Zheng, Wei-Shi},booktitle={Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},pages={30740--30749},month=jun,year={2026},}
Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning
ICCV
Zhi-Wei Xia , Kun-Yu Lin , Yuan-Ming Li , Wei-Jin Huang , Xian-Tuo Tan , and Wei-Shi Zheng
In Proceedings of the IEEE/CVF International Conference on Computer Vision, Oct 2025
@inproceedings{xia2025less,title={Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning},author={Xia, Zhi-Wei and Lin, Kun-Yu and Li, Yuan-Ming and Huang, Wei-Jin and Tan, Xian-Tuo and Zheng, Wei-Shi},booktitle={Proceedings of the IEEE/CVF International Conference on Computer Vision},pages={12894--12903},month=oct,year={2025},}
Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks
CVPR
Wei-Jin Huang , Yuan-Ming Li , Zhi-Wei Xia , Yu-Ming Tang , Kun-Yu Lin , Jian-Fang Hu , and Wei-Shi Zheng
In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Jun 2025
@inproceedings{huang2025modeling,title={Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks},author={Huang, Wei-Jin and Li, Yuan-Ming and Xia, Zhi-Wei and Tang, Yu-Ming and Lin, Kun-Yu and Hu, Jian-Fang and Zheng, Wei-Shi},booktitle={Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},pages={27794--27804},month=jun,year={2025},}
EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
ECCV
Yuan-Ming Li , Wei-Jin Huang(Co-First) , An-Lan Wang , Ling-An Zeng , Jing-Ke Meng , and Wei-Shi Zheng
In European Conference on Computer Vision, Jun 2024
@inproceedings{li2024egoexo,title={EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding},author={Li, Yuan-Ming and Huang, Wei-Jin and Wang, An-Lan and Zeng, Ling-An and Meng, Jing-Ke and Zheng, Wei-Shi},booktitle={European Conference on Computer Vision},pages={363--382},year={2024},organization={Springer},}
Education
Sun Yat-sen University (SYSU), Guangzhou, China2024.09 - Present
Ph.D., Computer Science and Technology
Specialization
Humanoid Learning
Video Action Understanding
South China University of Technology (SCUT), Guangzhou, China2020.09 - 2024.07