MasterI am a master's student majoring in Electronic Information (Artificial Intelligence) at South China University of Technology. Previously, I participated in a non-degree research program at the NUS (Chongqing) Research Institute, where I focused on reinforcement learning and multimodal large models.
My research interests center on Large Language Models and Agents. I have been fortunate to work on cutting-edge research projects, including AI-generated content detection and representation for multimodal generative tasks, with related work published in top-tier conferences such as AAAI and CVPR.
") does not match the recommended repository name for your site ("").
", so that your site can be accessed directly at "http://".
However, if the current repository name is intended, you can ignore this message by removing "{% include widgets/debug_repo_name.html %}" in index.html.
",
which does not match the baseurl ("") configured in _config.yml.
baseurl in _config.yml to "".

Xiaoye Zhu, Weixin Li, Junan Huo, Bozhong Wang, Jia Zeng, Yi Yang, Cen Chen, Qi Liu
ACM International Conference on Multimedia (ACM MM) 2026 CCF-A
We introduce CLARE, a clarification-aware and evolutionary 3D agent that resolves underspecified instructions through strategic dialogue before orchestrating 3D tools.
Xiaoye Zhu, Weixin Li, Junan Huo, Bozhong Wang, Jia Zeng, Yi Yang, Cen Chen, Qi Liu
ACM International Conference on Multimedia (ACM MM) 2026 CCF-A
We introduce CLARE, a clarification-aware and evolutionary 3D agent that resolves underspecified instructions through strategic dialogue before orchestrating 3D tools.

Jiaqi Chen*, Xiaoye Zhu*, Yue Wang*, Tianyang Liu, Xinhui Chen, Ying Chen, Chak Tou Leong, Yifei Ke, Joseph Liu, Yiwen Yuan, Julian McAuley, Li-jia Li (* equal contribution)
IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) 2025 CCF-A
We propose a symbolic generative task description language and inference engine, capable of representing arbitrary multimodal tasks as symbolic flows.
Jiaqi Chen*, Xiaoye Zhu*, Yue Wang*, Tianyang Liu, Xinhui Chen, Ying Chen, Chak Tou Leong, Yifei Ke, Joseph Liu, Yiwen Yuan, Julian McAuley, Li-jia Li (* equal contribution)
IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) 2025 CCF-A
We propose a symbolic generative task description language and inference engine, capable of representing arbitrary multimodal tasks as symbolic flows.

Jiaqi Chen*, Xiaoye Zhu*, Tianyang Liu*, Ying Chen, Xinhui Chen, Yiwen Yuan, Chak Tou Leong, Zuchao Li, Tang Long, Lei Zhang, Chenyu Yan, Guanghao Mei, Jie Zhang, Lefei Zhang (* equal contribution)
Association for the Advancement of Artificial Intelligence (AAAI) 2025 Oral CCF-A
Large Language Models (LLMs) have revolutionized text generation, making detecting machine-generated text increasingly challenging. Learn more
Jiaqi Chen*, Xiaoye Zhu*, Tianyang Liu*, Ying Chen, Xinhui Chen, Yiwen Yuan, Chak Tou Leong, Zuchao Li, Tang Long, Lei Zhang, Chenyu Yan, Guanghao Mei, Jie Zhang, Lefei Zhang (* equal contribution)
Association for the Advancement of Artificial Intelligence (AAAI) 2025 Oral CCF-A
Large Language Models (LLMs) have revolutionized text generation, making detecting machine-generated text increasingly challenging. Learn more