About

I am an Assistant Professor in the Computer Science Department at the College of Staten Island, City University of New York.

Before joining the City University of New York, I was a postdoctoral researcher at the University of Illinois Urbana-Champaign, where I was advised by Prof. Heng Ji and Prof. Chengxiang Zhai. I also worked with Prof. H. Chad Lane.

My research focuses on the knowledge lifecycle of large language models, with the goal of building more reliable and intelligent AI systems.

Recent Updates

[Aug 2026] Openings: I am looking for motivated students interested in interpretable, reliable, and trustworthy AI.
[Aug 2026] Our recent papers Geometric-disentanglement Unlearning (GU) and MemGuard were accepted to EMNLP 2026 Main; Current Agents Fail to Leverage World Model as Tool for Foresight to ACL 2026 Main; EMCompress to Findings of ACL 2026; ShortageSim to AAAI 2026 (Oral); ModelingAgent and SafeSwitch to Findings of EMNLP 2025; and How Do Analytic Design Choices Shape MLLM-Based Detection of Non-Verbal Behaviors in 360° Classroom Video to the EDM 2026 Proceedings!
[Aug 2026] I am joining the College of Staten Island, City University of New York as an Assistant Professor of Computer Science!
[Mar 2026] Call for papers for our 4th Towards Knowledgeable Foundation Models Workshop at ACL 2026.
[Mar 2026] Invited talk at the University of ArizonaKnowledge is Power, But Power Casts Shadows?
[Jan 2026] Invited talk at the University of EdinburghKnowledge is Power, But Power Casts Shadows?
[Jan 2026] Invited talk at the University of MassachusettsKnowledge is Power, But Power Casts Shadows?
[Jan 2026] Invited talk at the UIUC Large NLP Group SeminarKnowledge Lifecycle of LLMs
[Aug 2025] Organized the 3rd Towards Knowledgeable Foundation Models Workshop at ACL 2025.

Research Interest

My research spans NLP, LLM interpretation, and trustworthy AI, with a focus on how LLMs acquire, represent, and use knowledge to improve reliability and performance.

More broadly, I am interested in questions at the intersection of knowledge and reasoning: Can failures in AI reasoning be traced to deficiencies in how knowledge is represented and constructed? Can highly reliable reasoning emerge from better underlying knowledge building, updating, and organization?

My current research directions include:

  • Hallucination
    Understanding hallucination through the lens of knowledge representation and interaction, with the goal of interpreting, predicting, and preventing failures such as knowledge overshadowing.
  • Updating
    Developing reliable knowledge updating methods that incorporate new knowledge while preserving existing knowledge and model robustness.
  • Acquisition
    Improving knowledge acquisition and organization to enhance model intelligence and support stronger reasoning.

Selected Publications

see Google Scholar for all
Benchmarking Multi-turn Medical Diagnosis: Hold, Lure, and Self-correction
Jinrui Fang, Runhan Chen, Xu Yang, Jian Yu, Jiawei Xu, Ashwin Vinod, Wenqi Shi, Tianlong Chen, Heng Ji, Chengxiang Zhai, Ying Ding†, Yuji Zhang†
Preprintpaper
Atomic Reasoning for Scientific Table Claim Verification
Yuji Zhang, Qingyun Wang, Cheng Qian, Jiateng Liu, Chenkai Sun, Denghui Zhang, Tarek Abdelzaher, Chengxiang Zhai, Preslav Nakov, Heng Ji
Preprintpaper
The Law of Knowledge Overshadowing: Towards Understanding, Predicting, and Preventing LLM Hallucination
Yuji Zhang, Sha Li, Cheng Qian, Jiateng Liu, Pengfei Yu, Yi R. Fung, Chi Han, Kathleen McKeown, Chengxiang Zhai, Manling Li, Heng Ji
ACL 2025paper
Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models
Yuji Zhang, Sha Li, Jiateng Liu, Pengfei Yu, Yi R. Fung, Jing Li, Manling Li, Heng Ji
Preprintpaper
Geometric-disentanglement Unlearning
Duo Zhou*, Yuji Zhang*, Tianxin Wei, Ruizhong Qiu, Ke Yang, Xiao Lin, Cheng Qian, Jingrui He, Hanghang Tong, Heng Ji, Huan Zhang
EMNLP 2026paper
MemGuard: Preventing Memory Contamination in Long-Term Memory-Augmented Large Language Models
Hyeonjeong Ha, Jeonghwan Kim, Cheng Qian, Jiayu Liu, William M. Campbell, Yue Wu, Yuji Zhang, Kathleen McKeown, Dilek Hakkani-Tur, Heng Ji
EMNLP 2026paper
ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
Cheng Qian, Hongyi Du, Hongru Wang, Xiusi Chen, Yuji Zhang, Avirup Sil, Chengxiang Zhai, Kathleen McKeown, Heng Ji
EMNLP 2025paper
📢 ShortageSim: Simulating Drug Shortages under Information Asymmetry
Mingxuan Cui*, Yilan Jiang*, Duo Zhou*, Cheng Qian, Yuji Zhang†, Qiong Wang†
AAAI 2026 · Oralpaper

Invited Talks & Tutorials

  • Knowledge is Power, But Power Casts Shadows?
    City University of New York · May 2026
  • Knowledge is Power, But Power Casts Shadows?
    University of Arizona · Mar 2026
  • The Knowledge Lifecycle of LLMs: Memorization, Editing, and Beyond
    UIUC Large NLP Group Seminar · Jan 2026
  • Knowledge is Power, But Power Casts Shadows?
    University of Edinburgh · Jan 2026
  • Knowledge is Power, But Power Casts Shadows?
    University of Massachusetts · Jan 2026
  • Towards Knowledgeable Foundation Models
    NICE Academic Platform · Nov 2025
  • Robust and Trustworthy Language Models
    Northeastern University · Oct 2025
  • Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models
    Ploutos Academic Platform · May 2025
  • Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models
    The University of Texas at Austin · Mar 2025 · invited three-day academic visit
  • The Lifecycle of Knowledge in Large Language Models: Memorization, Editing, and Beyond
    AAAI 2025 Tutorial · Organizer and Speaker · Mar 2025
  • Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models
    Beijing Academy of Artificial Intelligence (BAAI) Academic Platform · Aug 2024

Teaching

  • Instructor · CSC 733: Natural Language Processing
    College of Staten Island, City University of New York · Fall 2026 · Mondays, 6:30 p.m.–9:10 p.m.
  • Guest Lecturer · CS 496: Agent AI
    Northwestern University · Spring 2026
  • Instructor · CS 591 BAI: Biologically Plausible Artificial Intelligence
    University of Illinois Urbana-Champaign · Spring 2026
  • Instructor · CS 591 BAI: Biologically Plausible Artificial Intelligence
    University of Illinois Urbana-Champaign · Fall 2025
  • Instructor · CS 591 BAI: Biologically Plausible Artificial Intelligence
    University of Illinois Urbana-Champaign · Spring 2025
  • Instructor · CS 591 BAI: Biologically Plausible Artificial Intelligence
    University of Illinois Urbana-Champaign · Fall 2024
  • Teaching Assistant · COMP 3134: Business Intelligence and CRM
    The Hong Kong Polytechnic University · Fall 2023
  • Teaching Assistant · COMP 5511: Artificial Intelligence Concepts
    The Hong Kong Polytechnic University · Spring 2023
  • Teaching Assistant · COMP 5511: Artificial Intelligence Concepts
    The Hong Kong Polytechnic University · Fall 2022
  • Teaching Assistant · COMP 4141: Crowdfunding and e-Finance
    The Hong Kong Polytechnic University · Spring 2022
  • Teaching Assistant · COMP 1433: Introduction to Data Analytics
    The Hong Kong Polytechnic University · Fall 2021
  • Teaching Assistant · COMP 1433: Introduction to Data Analytics
    The Hong Kong Polytechnic University · Spring 2021

Professional Service

  • Organizer · Workshop on Towards Knowledgeable Foundation Models
    ACL 2025; continuing workshop organization for ACL 2026
  • Session Chair · Language Models and Interpretability
    ACL 2025
  • Session Chair · Language Modeling
    ACL 2025
  • Area Chair
    ACL 2026 · LREC 2026
  • Conference Reviewer
    NeurIPS · ICLR · ACL · EMNLP · NAACL · KDD

Research Experience

University of Illinois Urbana-Champaign
Postdoctoral Researcher · 2024–2026
University of Illinois Urbana-Champaign
Visiting Ph.D. Student · 2023–2024
The Hong Kong Polytechnic University
Ph.D. in Computing · 2020–2024
Dartmouth College
Research Intern · 2019–2020