Shangjian Yin

Hi! My name is Shangjian Yin. I'm a CS PhD student at the University of California, Riverside, advised by Zhouxing Shi. Previously, I was a Research Intern at Microsoft AI and Meta AI.

Currently, I focus on LLM post-training, recursive self-improving LLMs, and long-horizon agents. I am also interested in world models, Physical AI, and agent systems infra, and I am open to collaboration!

Portrait of Shangjian Yin

Selected First-Author Research

Recursive Self-Improvement via On-Policy Distillation for Reasoning

Meta AI

GRLO: Towards Generalizable Reinforcement Learning in Open-Ended Environments from Zero

Under Review

Rethinking Reasoning Post-Training with General Chat Boosting and Dual-Reward Refinement

Under Review

From Individual to Common: An Early Exploration of Consensus in Non-verifiable Data for Preference Optimization

ACL 2026

Align Large Language Model with Human Preference via Extremely Self-Synthetic Data

ACL 2026

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

Microsoft AI

ECLM: Entity-Level Large Language Model for Spoken Language Understanding with Chain of Intent

ACL 2025

MIDLM: Multi-Intent Detection with Bidirectional Large Language Models

COLING 2025

Uni-MIS: United Multiple Intent Spoken Language Understanding via Multi-View Intent-Slot Interaction

AAAI 2024

Latest update: 09/2026.