I’m a full-time PhD candidate at HKUST Guangzhou, and I’m also an affiliate researcher at the IDEA Research, working on Large Language Models.
I am one of the earliest developers in China to conduct research on Pretrain Models, Vision-Language Model, and Diffusion-based AIGC model.
Internship and work experience include internet companies, state-owned enterprises, and research institutes.
Rich in technical passion and curiosity, enjoys sharing and exploring Data and Knowledge, Multimodal, Generative AI, and General Artificial Intelligence.
My current research focuses on LLM model training, Harness Engineering, datasets and data synthesis, and benchmark development, with an emphasis on Mathematical Modeling tasks.
News
[Jun. 2026] 🎉 Honored to be a co-author and contributor of SkillsBench, contributing tasks and BenchFlow SDK development.
[Apr. 2026] 🎉🎉 2 papers about LLM Data Engineering and Data Synthetic are accepted to ACL 2026 (1 Findings paper and 1 System Demostration paper).
[Nov. 2025] 🎉 Our paper about Anomaly Detection on Dynamic Graphs is accepted to AAAI 2026.
[Sept. 2025] 🎉 Our paper about Financial Large Language Models is accepted to EMNLP 2025 Findings.
[Jan. 2024] 🎉 Taiyi-Diffusion-XL Technical Report is released at Arxiv and HuggingFace
[Sept. 2023] 🎉 Ziya-VL Technical Report is released at Arxiv and HuggingFace and is reported by HuggingFace.