Policy Optimization for RL and LLM Alignment — Theory
The optimization story behind aligning language models with human preferences
I’m Ji Hun Wang, an Applied Scientist at Amazon working across research and engineering. My interests include post-training, interpretability, and AI safety and alignment. I’m also interested in formal accounts of natural language and the broader relationship between linguistic structure and computation.
Previously, I studied Computer Science and Linguistics at Stanford, completing B.A.S. and M.S. degrees.
The optimization story behind aligning language models with human preferences
About my favorite generative model of all time!
No essays yet.