I am currently a Ph.D. student at Sun Yat-sen University, specializing in Humanoid Learning and Video Understanding.

My research interests focus on humanoid learning, action recognition, and temporal video understanding. I am interested in building robust embodied and video understanding systems and aim to make meaningful contributions through sustained research.

Before starting my Ph.D., I was an undergraduate student at South China University of Technology studying Network Engineering. During my undergraduate studies, I was actively involved in algorithm competitions and research internships, gaining experience in computer vision and machine learning. I have received several honors, including the President Scholarship at Sun Yat-sen University, the National Scholarship, Kaggle Silver medals, and various competition prizes.

Feel free to reach out for research collaborations or academic discussions!

News

Jul 24, 2026 Our paper “ChoreoPlan: Hybrid Phrase Planning and Execution-Grounded Selection for Music-to-Humanoid Dance” was accepted to ACM MM 2026.
Feb 20, 2026 Our paper “Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations” was accepted to CVPR 2026.
Jun 25, 2025 Our paper “Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning” was accepted to ICCV 2025.
Feb 26, 2025 Our paper “Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks” was accepted to CVPR 2025.
Sep 06, 2024 Updated personal website content and contact email to huangwj235@mail2.sysu.edu.cn.

Selected Publications

  1. ChoreoPlan: Hybrid Phrase Planning and Execution-Grounded Selection for Music-to-Humanoid Dance
    ChoreoPlan: Hybrid Phrase Planning and Execution-Grounded Selection for Music-to-Humanoid Dance
    ACM MM
    Wei-Jin Huang , Jianhong Fan , Hao Huang , Zhi-Wei Xia , Jun-Yi Deng , Yuan-Ming Li , Kun-Yu Lin , and Wei-Shi Zheng
    In Proceedings of the ACM International Conference on Multimedia, 2026
  2. Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations
    Beyond Mimicry: Learning Whole-Body Human-Humanoid Interaction from Human-Human Demonstrations
    CVPR
    Wei-Jin Huang , Yue-Yi Zhang , Yi-Lin Wei , Zhi-Wei Xia , Juantao Tan , Yuan-Ming Li , Zhilin Zhao , and Wei-Shi Zheng
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Jun 2026
  3. Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning
    Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning
    ICCV
    Zhi-Wei Xia , Kun-Yu Lin , Yuan-Ming Li , Wei-Jin Huang , Xian-Tuo Tan , and Wei-Shi Zheng
    In Proceedings of the IEEE/CVF International Conference on Computer Vision, Oct 2025
  4. Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks
    Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks
    CVPR
    Wei-Jin Huang , Yuan-Ming Li , Zhi-Wei Xia , Yu-Ming Tang , Kun-Yu Lin , Jian-Fang Hu , and Wei-Shi Zheng
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Jun 2025
  5. EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
    EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
    ECCV
    Yuan-Ming Li , Wei-Jin Huang (Co-First) , An-Lan Wang , Ling-An Zeng , Jing-Ke Meng , and Wei-Shi Zheng
    In European Conference on Computer Vision, Jun 2024

Education

Sun Yat-sen University (SYSU), Guangzhou, China 2024.09 - Present
Ph.D., Computer Science and Technology
  • Specialization
    • Humanoid Learning
    • Video Action Understanding
South China University of Technology (SCUT), Guangzhou, China 2020.09 - 2024.07
B.E., Network Engineering
  • Average grade ranking 3/66