Sijia L. Liu

Email  /  Google Scholar  /  Twitter  /  Github

sijial_pic_cropped.jpg

sijia.liu@{myschool}.edu

35 Olden St

Princeton, NJ 08540

PhD Student of Computer Science at Princeton University

Research Areas: ML / NLP / Reasoning

Hello! I am a second-year CS PhD student at Princeton University, advised by Prof. Karthik Narasimhan at Princeton Language and Intelligence. My research primarily focuses on LLM post-training and long-horizon agents. I will also be joining Allen Institute for Artificial Intelligence (AI2) as a summer Research Intern working on long-horizon hypothesis generation. Previously, I spent three years as a research scientist on the Amazon AGI Post-Training team, contributing to the development of Nova — state-of-the-art text and multimodal models and the Alexa Prize SocialBot, working with Dr. Yang Liu and Prof. Dilek Hakkani-Tur in Sunnyvale, CA. Before that, I received my bachelor’s from Peking University and my master’s from Carnegie Mellon University.

Outside of work, I value health and enjoy sports including snowboarding, badminton, and running (half marathon PR: 02:10:15, Brooklyn, Apr 2025). Being a language lover at heart, I’ve been pre-trained on Chinese and fine-tuned on English, Korean, German, and Japanese — though I might suffer from catastrophic forgetting sometimes. I am also a cat lover and own a cute beige American shorthair.


news

Jun 15, 2026 We introduce ContextRL, a context-aware RL objective that rewards models for selecting the context that supports an answer, improving long-horizon agentic and multimodal reasoning by +2.2% and +1.8% over standard GRPO.
Jan 26, 2026 Humanline has been accepted to ICLR 2026. See you in Rio!
Dec 12, 2025 Received the generous Tinker teaching grant to support our undergraduate Introduction to NLP course.
Dec 11, 2025 Received the Tinker research grant ($5,000) to support our research on post-training algorithms.
Oct 3, 2025 We introduce Humanline as a simple yet surprisingly effective design to close the online-offline alignment gap across multiple model families (Llama-3, Gemma-2) and model sizes (1.5B-27B) on both instruction-following and mathematical reasoning.

selected publications

  1. contextrl-overview.png
    Context-Aware RL for Agentic and Multimodal LLMs
    Peiyang Xu, Bangzheng Li, Sijia Liu, and 4 more authors
    2026
  2. humanline-perception.png
    Humanline: Online Alignment as Perceptual Loss
    Sijia Liu, Niklas Muennighoff, and Kawin Ethayarajh
    2025
  3. nova-family.png
    The Amazon Nova family of models: Technical report and model card
    Amazon Artificial General Intelligence
    Amazon Technical Reports, 2024
  4. decrim-pipeline.png
    LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
    Thomas Palmeira Ferraz, Kartik Mehta, Yu-Hsiang Lin, and 7 more authors
    EMNLP, NeurIPS Workshop on System 2 Reasoning at Scale (Oral), 2024
  5. interactive-eval.png
    Towards Credible Human Evaluation of Open-Domain Dialog Systems Using Interactive Setup
    Sijia Liu, Patrick Lange, Behnam Hedayatnia, and 5 more authors
    AAAI (Oral), NeurIPS Workshop on Human Evaluation of Generative Models (Oral), EMNLP Workshop on Natural Language Generation, Evaluation, and Metrics (Oral), Jun 2023
  6. contradiction-rewriting.png
    Improving Bot Response Contradiction Detection via Utterance Rewriting
    Di Jin, Sijia Liu, Yang Liu, and 1 more author
    SIGDIAL (Oral), Jul 2022