I am currently a Ph.D. candidate in Computer Science and Technology at the University of Cambridge, under the supervision of Prof. Rafal Mantiuk. I obtained my Bachelor of Engineering degree from Fudan University in 2023. In 2022, I had the opportunity to work as a research assistant at Stanford University - SVL , advised by Prof. Jiajun Wu and Prof. Yunzhu Li. Prior to that, I served as a research assistant at Fudan University from 2021 to 2023, under the guidance of Prof. Tao Chen. More recently, I worked as a Research Scientist Intern building Vidi-Gen at ByteDance USA, and with the Computational Photography team at LG Electronics USA.
I am highly interested in the fields of computer vision, computer graphics, multimodal LLMs, and video generation. My doctoral research bridges insights from the human visual system with cutting-edge machine learning techniques, establishing robust computational models for the evaluation of display image and video quality. More recently, this interest has extended into multimodal large language models and long-form video generation, an area I am now actively working on.
Email: yc613 [at] cam [dot] ac [dot] uk & cycxueshu [at] 163 [dot] com
Wechat: cyc13700232963
OpenReview /
GitHub /
知乎Zhihu /