Yixiong Hao

Yixiong Hao

Hi, I'm Yixiong!

I study CS at Georgia Tech. I work on technical research and strategy to reduce societal scale catastrophic risks from highly general, increasingly superintelligent AI systems.

I'm currently a grantmaker at The Astralis Foundation through the Astra Fellowship, where I will be focusing on macro AI strategy and capacity building in Asia. I'm also a co-founder of Second Look Research, and an advisor to the Georgia Tech AI Safety Initiative.

Previously, I

Outside of work, I generally love being active. I used to play competitive golf and table tennis. I also like checking out new cafes and restaurants!

Check out some of my amazing friends, collaborators, and mentors. Parv, Will, Zephy, Jasmine, Jo, Andrew, Andy, Rocio, Tzu, Mantas, Animesh, Kartik, Abdur Raheem, Max.

Research

I'm interested in understanding what and how general AI systems learn & generalize in order to develop scalable methods for alignment, evaluations, and control. My latest projects are to do with agentic misalignment in RL with CAIS, and interpretability of VLA models with the PAIR Lab. Research blogs and unpublished work are under writing.

A. Ghosh, A. Reuel, J. Chim, …, Y. Hao, …, I. Solaiman
arXiv preprint 2026
A. K. Saeri, J. Graham, M. Noetel, …, Y. Hao, …, N. Thompson
arXiv preprint 2026
Y. Hao, A. Panda, S. Shabalin, S. A. R. Ali
ICLR 2025 Human-AI Co-evolution Workshop + 2 others
S. Shabalin, A. Panda, D. Kharlapenko, A. R. Ali, Y. Hao, A. Conmy
CVPR 2025 XIV Workshop
G. Matlin, P. Mahajan, I. Song, Y. Hao, R. Bard, S. Topp, E. Montoya, M. R. Parwani, S. Shetty, M. Riedl
EMNLP 2025 Wordplay Workshop