Text-to-Image & Video
Generating high-quality visual content with stronger semantic alignment and cross-frame consistency.
Multimodal Generative AI · LLMs
林哲栋
Hi, I’m Zhedong Lin (林哲栋). I am currently preparing to begin my PhD studies at the University of Auckland, where I will continue my research under the supervision of Prof. Jiamou Liu and Prof. Xinyu Zhang. I hold an MSc in Artificial Intelligence from the University of Auckland and received my BSc in Computer Science and Technology from Southwest University in 2025. My research interests include large language models (LLMs), multimodal generative AI, and their applications in text-to-image synthesis, text-to-video generation, image and video editing, and controllable diffusion models. If you are interested in my research or would like to discuss potential collaborations, please feel free to contact me at zlin629@aucklanduni.ac.nz.

Generating high-quality visual content with stronger semantic alignment and cross-frame consistency.
Enabling flexible, user-guided editing while preserving context, identity, and visual coherence.
Improving fine-grained control over generative processes for creative and practical applications.
Studying how LLMs and multimodal models support generation, editing, and cross-modal understanding.
Zhongsheng Wang, Zhedong Lin, Qian Liu, Xinyu Zhang, Jiamou Liu
ACM Multimedia 2026 · CCF-A
Zhedong Lin, Zhongsheng Wang, Qian Liu, Xinyu Zhang, Jiamou Liu
Artificial Intelligence Review · JCR Q1, SCI Q1 Top
Zhongsheng Wang, Ming Lin, Zhedong Lin, Yaser Shakib, Qian Liu, Jiamou Liu
ACM Multimedia Asia 2025 · CCF-C
My academic path from Southwest University to postgraduate research at the University of Auckland.
Academic background →Teaching assistance in advanced machine learning and mathematics for computer science, plus a guest lecture on Agentic AI.
Teaching experience →Research, scholarship, course, and graduation distinctions across my academic journey.
View honors →Code Vibe Reading is a VS Code extension for AI-assisted code understanding and navigation.
View project →