About Me

I am a Postdoctoral Scholar in the Computer Science Department at Stanford University, where I have the privilege of being advised by Prof. Sanmi Koyejo, in the Stanford Trustworthy AI Research (STAIR) Lab.

Research Interests

I strive to advance trustworthy and responsible AI. My research spans agentic systems safety enhancement, in-situ behavioral evaluation of generative AI and agentic systems, causal learning and reasoning to facilitate and enhance the capacity of intelligent systems, and ML fairness and computational justice. My ultimate goal is to cultivate safe and principled intelligence, so that technology can improve our lives with transparent responsibility and clear purpose. I seek to foster a symbiotic dance between artificial and natural intelligence, where they inspire, collaborate, and enhance each other to drive scientific discovery and support societal progress.

News

July 2026 Our paper “In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores” is accepted to COLM 2026. We argue that LLM fairness should be evaluated through in-situ behavioral pattern rather than standardized-test Q&A benchmarks.
May 2026 Our position paper “Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants” is accepted to ICML 2026 Position Paper Track. We argue that ML fairness should quantify structural injustice via social determinants, beyond sensitive attributes.
April 2026 We are organizing the Algorithmic Fairness Across Alignment Procedures and Agentic Systems (AFAA) Workshop at ICLR 2026, April 26, 2026, in Rio de Janeiro, Brazil!

Selected Publications

* denotes equal contribution

  1. COLM
    In-Situ Behavioral Evaluation for LLM Fairness, Not Standardized-Test Scores
    In Third Conference on Language Modeling, 2026.
  2. ICMLPosition
    Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinants
    In Forty-Third International Conference on Machine Learning, 2026.
  3. Reflection-Window Decoding: Text Generation with Selective Refinement
    In Forty-Second International Conference on Machine Learning, 2025.
  4. Prompting Fairness: Integrating Causality to Debias Large Language Models
    In Thirteenth International Conference on Learning Representations (preliminary version titled "Steering LLMs Towards Unbiased Responses: A Causality-Guided Debiasing Framework"), 2025.
  5. ICLRSpotlight
    Procedural Fairness Through Decoupling Objectionable Data Generating Components
    Zeyu TangJialu WangYang LiuPeter Spirtes, and Kun Zhang
    In Twelfth International Conference on Learning Representations (preliminary version presented in NeurIPS 2023 AFT workshop), 2024.
  6. What-is and How-to for Fairness in Machine Learning: A Survey, Reflection, and Perspective
    Zeyu TangJiji Zhang, and Kun Zhang
    ACM Computing Surveys, 2023.
  7. Tier Balancing: Towards Dynamic Fairness over Underlying Causal Factors
    Zeyu TangYatong ChenYang Liu, and Kun Zhang
    In Eleventh International Conference on Learning Representations (preliminary version presented in NeurIPS 2022 AFCP workshop), 2023.
  8. CLeaRSpotlight
    Attainability and Optimality: The Equalized Odds Fairness Revisited
    Zeyu Tang, and Kun Zhang
    In First Conference on Causal Learning and Reasoning, 2022.