About Me
I am currently a 2nd-year Ph.D. student at the University of Science and Technology of China (USTC), fortunate to be co-advised by Prof. Zheng-Jun Zha and Associate Prof. Jiawei Liu. Previously, I received my Bachelor’s degree in Automation from Northeastern University (2019-2023).
Research Interests:
- Multimodal Reasoning and Understanding - Developing models capable of understanding and reasoning over complex multimodal and long-horizon information
- Self-Evolving Agents - Enabling agents to continuously acquire, refine, and internalize reusable knowledge and strategies through interaction and experience
- Agentic Reinforcement Learning — Training general / domain-specific agents with reinforcement learning for stronger reasoning, tool use, and long-horizon decision-making
🔥 News
- 2026.04: 📚 One paper on VLM-based class incremental learning (third author) was accepted by CVM 2026.
- 2026.02: 🎉 One paper on open-vocabulary HOI detection (first author) was accepted by CVPR 2026 (highlight).
- 2026.01: 📚 One paper on active prompt learning (fifth author) was accepted by ICLR 2026.
- 2025.11: 🎉 One paper on zero-shot HOI detection (co-first author) was accepted by IJCV 2026.
- 2025.10: 📚 One paper on active prompt learning (fourth author) was accepted by IJCV 2026.
- 2025.07: 🏆 One paper was accepted by MM 2025 Workshop and received Third Place.
- 2025.02: 📚 One paper on multi-task test-time adaptation (fifth author) was accepted by CVPR 2025.
- 2024.12: 🎉 One paper on HOI detection (first author) was accepted by AAAI 2025.
- 2024.07: 🏆 One paper was accepted by MM 2024 Workshop and received Second Place.
- 2024.01: 📚 One paper on multimodal fact-checking (fifth author) was accepted by WWW 2024.
📖 Education
- 2025.09 - Present |
Ph.D. in Electronic and Information Engineering
University of Science and Technology of China - 2023.09 - 2025.06 |
M.Eng. in Electronic and Information Engineering
University of Science and Technology of China - 2019.09 - 2023.06 |
B.Eng. in Automation
Northeastern University
📝 Publications

Mining the Potential of Rehearsal Mechanism for VLM-based Class Incremental Learning
Sen Tao, Jiawei Liu, Yongchao Xu, Guangxi Wan, Peng Zeng.
Chinese Conference on Computer Vision and Machine Intelligence (CVM), 2026.
PDF Code
We propose a debiased memory-calibrated Gaussian discriminant analysis method for VLM-based class incremental learning.

Learning to Diversify and Focus: A Reinforcement Framework for Open-Vocabulary HOI Detection
Yongchao Xu, Jiawei Liu, Junfeng Wang, Sen Tao, Na Jiang, Zheng-Jun Zha.
Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), 2026.
PDF Supp Code
We propose a semantic-diversified and interaction-focused framework for open-vocabulary HOI detection.

PSP: Prompt-Guided Self-Training Sampling Policy for Active Prompt Learning
Sen Tao, Kaiduo Feng, Jiawei Liu, Peng Zeng, Yongchao Xu, Yufei Zheng, Zheng-Jun Zha.
International Conference on Learning Representations (ICLR), 2026.
PDF Code
We propose a prompt-guided self-training sampling policy for active prompt learning.

Mamba-Driven Comprehensive Context Learning for Zero-Shot HOI Detection
Jiawei Liu, Yongchao Xu (co-first author), Sen Tao, Yuexuan Qi, Zheng-Jun Zha.
International Journal of Computer Vision (IJCV), 2026.
PDF Code
We propose a comprehensive context learning framework for zero-shot HOI detection.

Boosting Active Prompt Learning via Discriminative Self-Training Dual-Curriculum Learning
Sen Tao, Jiawei Liu, Peng Zeng, Yongchao Xu, Bingyu Hu, Zheng-Jun Zha.
International Journal of Computer Vision (IJCV), 2026.
PDF Code
We propose a dual-curriculum active prompt learning framework.

HOIMamba: Efficient Mamba-based Disentangled Progressive Learning for HOI Detection
Yongchao Xu, Jiawei Liu, Sen Tao, Qiang Zhang, Zheng-Jun Zha.
The Thirty-Ninth AAAI Conference on Artificial Intelligence (AAAI), 2025.
PDF Code
We introduce a Mamba-based decoder for robust HOI detection.

Hierarchical Knowledge Prompt Tuning for Multi-task Test-Time Adaptation
Qiang Zhang, Mengsheng Zhao, Jiawei Liu, Fanrui Zhang, Yongchao Xu, Zheng-Jun Zha.
Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), 2025.
PDF Code
We address multi-task test-time adaptation for pre-trained VLMs.

ESCNet: Entity-enhanced and Stance Checking Network for Multimodal Fact-Checking
Fanrui Zhang, Jiawei Liu, Jingyi Xie, Qiang Zhang, Yongchao Xu, Zheng-Jun Zha.
International World Wide Web Conference (WWW), 2024.
PDF Code
We establish a large-scale, multi-domain Chinese multimodal fact-checking dataset.
🏆 Honors & Awards
- 2026.09: First-Class Scholarship of USTC (校一等奖学金).
- 2025.09: First-Class Scholarship of USTC (校一等奖学金).
- 2024.09: First-Class Scholarship of USTC (校一等奖学金).
- 2023.09: First-Class Scholarship of USTC (校一等奖学金).
- 2023.06: Outstanding Graduate of Liaoning Province (辽宁省优秀毕业生,Top 1%).
- 2023.05: Principal’s Medal of NEU (东北大学校长奖章, 全校10人).
- 2022.09: First-Class Scholarship of NEU (校一等奖学金,Top 3%).
- 2022.05: Outstanding Winner and AMS Award of the International Collegiate Mathematical Modeling Competition (美赛O奖&AMS奖,Top 0.02%).
- 2020.09: National Scholarship of China (本科生国家奖学金, Top 1%); First-Class Scholarship of NEU (校一等奖学金,Top 3%).