I am a 2nd year PhD student at Carnegie Mellon University’s Language Technologies Institute, co-advised by Professor Mona Diab at CMU and Professor Zhijing Jin at the University of Toronto. My research focuses on reasoning, continual learning, reinforcement learning, agents, and alignment for large language models.
I am currently a research intern at Meta in Redmond, hosted by Xin Luna Dong, Renjie Tao, and Tony Liao, working on accelerating ML research through harnesses and RL for research ideation. Previously, I interned at Amazon Rufus in Seattle in 2025, where I worked on honesty alignment for reasoning models; at Amazon AWS Bedrock in New York in 2024, where I studied multimodal hallucination; and at WarpEngine in Shanghai in 2023, where I built a personalized chatbot product.
At CMU, I am a teaching assistant for 11-768 AI Agents (Fall 2026, with Graham Neubig and Daniel Fried) and was a teaching assistant for 11-830 Ethics, Safety, and Social Impact in NLP and LLMs (Spring 2026, with Maarten Sap). I co-organized the PersonaNLP workshop at NeurIPS and the NLP for Positive Impact workshop in 2025, and I review for NeurIPS, ICML, ICLR, COLM, ACL ARR, and CHI.
Before CMU, I worked with Professors Rada Mihalcea and Zhijing Jin in the LIT group at the University of Michigan on NLP for social good. I have also had a great time working with Professor Wei Hu on improving worst-group robustness.
PhD in Language Technologies, Aug 2025 - May 2028 (Expected)
Carnegie Mellon University, Pittsburgh
Master of Language Technologies, Aug 2023 - May 2025
Carnegie Mellon University, Pittsburgh
BSE in Computer Science, Sept 2021 - May 2023
University of Michigan, Ann Arbor
BSE in Electrical and Computer Engineering, Sept 2019 - Aug 2023
Shanghai Jiao Tong University (UM-SJTU Joint Institute)
"Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases"
Preprint 2026
"PaperMentor: A human-centered multi-agent writing tutor for AI research papers in Overleaf"
ACL 2026 Demo
"Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR"
Best Paper Award at ICML 2026 RLxF Workshop
"Self-distillation zero: Self-revision turns binary rewards into dense supervision"
Best Paper Award at ICML 2026 RLxF Workshop
"CLT-Forge: A Scalable Library for Cross-Layer Transcoders and Attribution Graphs"
Preprint 2026
"Making Complex Reasoning Student-Friendly: A Hybrid LLM-to-SLM Distillation Framework"
Preprint 2026
"Stabilizing Reinforcement Learning for Honesty Alignment in Language Models on Deductive Reasoning"
AAAI 2026 Bridge LMReasoning Workshop, AAAI 2026 MATH4AI Workshop
"LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Use"
COLM 2025 Interplay Workshop
"CauSciBench: A Comprehensive Benchmark on End-to-End Causal Inference for Scientific Research"
ICML 2026
"Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics"
EMNLP 2025 Main
"Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design"
EMNLP 2025 Main Oral
"BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data"
ACL 2025 Main
"Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)"
ACL 2025 Findings
"Uncovering and Understanding Social Media Censorship across Countries"
ACL 2025 Findings
"Language Model Alignment in Multilingual Trolley Problems"
ICLR 2025 and Best Paper Award at NeurIPS 2024 Pluralistic Alignment Workshop
"Synatra: Turning indirect knowledge into direct demonstrations for digital agents at scale"
NeurIPS 2024
"Inducing Elasticity in Foundation Models: Post-Training Techniques for Adaptable Inference"
NeurIPS ENLSP Workshop 2024
"Chumor 1.0: A Truly Funny and Challenging Chinese Humor Understanding Dataset from Ruo Zhi Ba"
Preprint 2024
"Voices of Her: Analyzing Gender Differences in the AI Publication World"
ACL 2025 NLP for Positive Impact Workshop