I study how multimodal agents turn visual experience into grounded reasoning and purposeful action. I build systems that understand long-horizon visual experiences, use tools to seek and verify evidence, and adapt across physical and digital environments.
Shulin Tian is a Ph.D. student in Computer Science at Nanyang Technological University, advised by Prof. Ziwei Liu and Dr. Hongyuan Zhu. Her research centers on multimodal agents and visual reasoning, with an emphasis on long-horizon multimodal understanding and tool-augmented reasoning. Previously, she worked with Prof. Ranjay Krishna at the University of Washington on vision-language reasoning, and with Prof. Bihan Wen at NTU on computational imaging. She received her B.Eng. in Electrical and Electronic Engineering with Honours (Highest Distinction) from NTU and is supported by the A*STAR Computing and Information Science Scholarship.
News
08/2026V-Rubrics is accepted to EMNLP 2026 Main; Demo-ICL is accepted to Findings.
08/2026We release Open Evaluation Agent including EA-CoT-10K and EA-3B for open-source, promptable evaluation of visual generative models.
06/2026We release S-Agent, a spatial tool-use agent for continuous multi-view image and video reasoning.