Chuyan Zhou

Researching on [Alignment of/Agentic/Diffusion/Scaling] LLMs
Japanese C1, English C1, spoken & written
Feel free to contact for collaboration
Click for my Curriculum Vitae

Research Interests

  • Diffusion LLMs: Architecture, Agentic Post-training/Scaling
  • Agents: Agentic RL infrastructure and algorithms, Harness Design, Recursive Self-Improvement
  • Alignment (Post-training) & Nonparametric Scaling of LLMs: Reinforcement Finetuning, Process Reward Models & Design; Distillation

Nonzero Experience / Secondary Interests

Underlined parts are with nonzero experiences.

  • LLM Efficiency: Looped Transformers, Soft/Latent Token Transformers, Speculative/Jacobi Decoding (interested in dLLM-powered Decoding Algos), Quantization, KV Cache Compression/Pruning
  • Mechanistic Interpretability: Sparse Autoencoders, Steering, Logit/Jacobian Lens
  • Combination of Connectionist and Probabilistic/Symbolic Methods for NLP
  • AI4S: Neural Methodologies for Biology/Physics/Chemistry/…

Publications

GiLT: Augmenting Transformer Language Models with Dependency Graphs

ACL 2026 Main Conference
Tianyu Huang, Yida Zhao, Chuyan Zhou, Kewei Tu [Paper]

History

  • 2026.09.28: Joined Okazaki Lab officially along with formal enrollment.
  • 2026.07.15: Graduated from ShanghaiTech University with a bachelor degree in Computer Science as a distinguished alumnus.
  • 2026.05.28: Received an offer of admission to 東京科学大学大学院情報理工学院 情報工学系 知能情報コース (Major of AI, Department of CS, School of Computing, ISCT) for M.S. in the Okazaki Laboratory.
  • 2026.04.07: GiLT: Augmenting Transformer Language Models with Dependency Graphs with my participation, was accepted to the ACL 2026 Main Conference.
  • 2026.03.06: Received an offer of admission to the master’s program in 東京大学大学院 工学系研究科 技術経営戦略学専攻 (International Technology Management, G30-TMI) at UT, with placement in my first-choice Matsuo-Iwasawa Laboratory.
  • 2026.02.13: Received an offer of admission to the master’s program in 東京大学大学院情報理工学系研究科 創造情報学専攻 (Creative Informatics at the University of Tokyo), with placement in my first-choice Nakayama Laboratory.
  • 2025.05: Received the Outstanding Student Award and Scholarship from ShanghaiTech University.
  • 2025.01: Completed the GLOBE Program in University of California Berkeley starting from 2024.08, majoring in Computer Science, GPA 4.0/4.0.
  • 2023.10: Won a silver medal (24th place) teamed in Kaggle Bengali.AI Speech Recognition Challenge.
  • 2023.10: Joined Prof. Kewei Tu’s research group at ShanghaiTech University as an undergraduate researcher.

Recent Posts

My TLDR for Tsubame 4 @ISCT

A TLDR-like manual of Tsubame (Altair) tailored only to myself being used to Slurm.
2026-10-10
1 min read

MDP & Timestep RL for Discrete Diffusion LLMs

A primitive formulation of dLLM RL with denoising timesteps as actions.
2026-09-30
6 min read
Featured Image

Training-free Verifiable Process Reward for LLM Reinforcement Finetuning

A training-free process reward system based on formalized theorem proving for LLM reinforcement finetuning.
2026-06-15
1 min read