Jiarui Liu

Jiarui Liu

PhD Student in Language Technologies

Carnegie Mellon University, Pittsburgh

Biography

I am a 2nd year PhD student at Carnegie Mellon University’s Language Technologies Institute, co-advised by Professor Mona Diab at CMU and Professor Zhijing Jin at the University of Toronto. My research focuses on reasoning, continual learning, reinforcement learning, agents, and alignment for large language models.

I am currently a research intern at Meta in Redmond, hosted by Xin Luna Dong, Renjie Tao, and Tony Liao, working on accelerating ML research through harnesses and RL for research ideation. Previously, I interned at Amazon Rufus in Seattle in 2025, where I worked on honesty alignment for reasoning models; at Amazon AWS Bedrock in New York in 2024, where I studied multimodal hallucination; and at WarpEngine in Shanghai in 2023, where I built a personalized chatbot product.

At CMU, I am a teaching assistant for 11-768 AI Agents (Fall 2026, with Graham Neubig and Daniel Fried) and was a teaching assistant for 11-830 Ethics, Safety, and Social Impact in NLP and LLMs (Spring 2026, with Maarten Sap). I co-organized the PersonaNLP workshop at NeurIPS and the NLP for Positive Impact workshop in 2025, and I review for NeurIPS, ICML, ICLR, COLM, ACL ARR, and CHI.

Before CMU, I worked with Professors Rada Mihalcea and Zhijing Jin in the LIT group at the University of Michigan on NLP for social good. I have also had a great time working with Professor Wei Hu on improving worst-group robustness.

Interests
  • Reasoning
  • Continual learning
  • Reinforcement learning
  • Agents
  • Alignment
Education
  • PhD in Language Technologies, Aug 2025 - May 2028 (Expected)

    Carnegie Mellon University, Pittsburgh

  • Master of Language Technologies, Aug 2023 - May 2025

    Carnegie Mellon University, Pittsburgh

  • BSE in Computer Science, Sept 2021 - May 2023

    University of Michigan, Ann Arbor

  • BSE in Electrical and Computer Engineering, Sept 2019 - Aug 2023

    Shanghai Jiao Tong University (UM-SJTU Joint Institute)

News📢

  • [Aug. 2026] I am a TA for 11-768 AI Agents at CMU this fall!
  • [Jul. 2026] Two papers received Best Paper Awards at the ICML 2026 RLxF Workshop!
  • [May. 2026] I started my research intern at Meta in Redmond!
  • [May. 2026] PaperMentor to appear in ACL 2026 Demo and CauSciBench to appear in ICML 2026!
  • [Jan. 2026] Two papers to appear in EACL 2026!
  • [Jan. 2026] I am a TA for 11-830 Ethics, Safety, and Social Impact in NLP and LLMs at CMU this spring!
  • [Dec. 2025] Invited talk at Nice-NLP on honest language models for deductive reasoning.
  • [Nov. 2025] Two papers to appear in EMNLP 2025, see you in Suzhou!
  • [May. 2025] I started my applied scientist intern at Amazon Rufus in Seattle!
  • [May. 2025] Four papers to appear in ACL 2025!
  • [Jan. 2025] One paper to appear in ICLR 2025!
  • [Dec. 2024] Best Paper Award at the NeurIPS 2024 Pluralistic Alignment Workshop!
  • [Sep. 2024] One paper to appear in EMNLP 2024, see you in Miami!
  • [May. 2024] I started my applied scientist intern at Amazon AWS!
  • [Mar. 2024] One paper to appear in NAACL 2024 as Oral Presentation!
  • [Jan. 2024] One paper to appear in ICLR 2024, see you in Vienna!

Publications

* indicates equal contribution.

"Learning from Use: Test-Time Learning in Large Language Models and Agents"


Preprint 2026

"Evaluating Multi-Turn Multimodal Diagnostic Reasoning on Challenging Real-World Clinical Cases"


Preprint 2026

"PACE: A Proxy for Agentic Capability Evaluation"


Preprint 2026

"PaperMentor: A human-centered multi-agent writing tutor for AI research papers in Overleaf"


ACL 2026 Demo

"OdysSim: Building Foundation Models for Human Behavior Simulation"


Preprint 2026

"Re-Centering Humans in LLM Personalization"


Preprint 2026

"Knowledge Index of Noah's Ark"


Preprint 2026

"LT2: Linear-Time Looped Transformers"


Preprint 2026

"Reinforcing Human Behavior Simulation via Verbal Feedback"


Preprint 2026

"MixSD: Mixed Contextual Self-Distillation for Knowledge Injection"


Preprint 2026

"Transferability for General Reasoning: An Automated Curriculum for Multi-Domain RLVR"


Best Paper Award at ICML 2026 RLxF Workshop

"Self-distillation zero: Self-revision turns binary rewards into dense supervision"


Best Paper Award at ICML 2026 RLxF Workshop

"CLT-Forge: A Scalable Library for Cross-Layer Transcoders and Attribution Graphs"


Preprint 2026

"Mind the sim2real gap in user simulation for agentic tasks"


Preprint 2026

"Making Complex Reasoning Student-Friendly: A Hybrid LLM-to-SLM Distillation Framework"


Preprint 2026

"Stabilizing Reinforcement Learning for Honesty Alignment in Language Models on Deductive Reasoning"


AAAI 2026 Bridge LMReasoning Workshop, AAAI 2026 MATH4AI Workshop

"LLM Microscope: What Model Internals Reveal About Answer Correctness and Context Use"


COLM 2025 Interplay Workshop

"CauSciBench: A Comprehensive Benchmark on End-to-End Causal Inference for Scientific Research"


ICML 2026

"Taming Object Hallucinations with Verified Atomic Confidence Estimation"


EACL 2026

"CORE: Measuring Multi-Agent LLM Interaction Quality under Game-Theoretic Pressures"


EACL 2026

"Synthetic Socratic Debates: Examining Persona Effects on Moral Decision and Persuasion Dynamics"


EMNLP 2025 Main

"Humanizing Machines: Rethinking LLM Anthropomorphism Through a Multi-Level Framework of Design"


EMNLP 2025 Main Oral

"BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data"


ACL 2025 Main

"Towards Global AI Inclusivity: A Large-Scale Multilingual Terminology Dataset (GIST)"


ACL 2025 Findings

"Uncovering and Understanding Social Media Censorship across Countries"


ACL 2025 Findings

"Chumor 2.0: Towards Benchmarking Chinese Humor Understanding"


ACL 2025 Findings

"EmoNews: A Spoken Dialogue System for Expressive News Conversations"


SigDial 2025 Demo

"Language Model Alignment in Multilingual Trolley Problems"


ICLR 2025 and Best Paper Award at NeurIPS 2024 Pluralistic Alignment Workshop

"Implicit Personalization in Language Models: A Systematic Study"


EMNLP 2024 Findings

"Synatra: Turning indirect knowledge into direct demonstrations for digital agents at scale"


NeurIPS 2024

"Inducing Elasticity in Foundation Models: Post-Training Techniques for Adaptable Inference"


NeurIPS ENLSP Workshop 2024

"Chumor 1.0: A Truly Funny and Challenging Chinese Humor Understanding Dataset from Ruo Zhi Ba"


Preprint 2024

"Automatic Generation of Model and Data Cards: A Step Towards Responsible AI"


NAACL 2024 Oral

"Analyzing the Role of Semantic Representations in the Era of Large Language Models"


NAACL 2024

"Can Large Language Models Infer Causation from Correlation?"


ICLR 2024

"Bias Amplification Enhances Minority Group Performance"


TMLR 2023

"Voices of Her: Analyzing Gender Differences in the AI Publication World"


ACL 2025 NLP for Positive Impact Workshop