Yixiong Hao郝奕雄

Yixiong Hao

Hi, I'm Yixiong!

你好,我是奕雄!

I study CS at Georgia Tech. I work on technical research and strategy to reduce societal scale catastrophic risks from highly general, increasingly superintelligent AI systems.

我在佐治亚理工学院(Georgia Tech)学习计算机科学。我从事技术研究与战略工作,致力于降低高度通用、日益趋近超级智能的 AI 系统所带来的社会规模灾难性风险

I'm currently a grantmaker at The Astralis Foundation through the Astra Fellowship, where I will be focusing on macro AI strategy and capacity building in Asia. I'm also a co-founder of Second Look Research, and an advisor to the Georgia Tech AI Safety Initiative.

我目前通过 Astra Fellowship 在 Astralis Foundation 担任资助官(grantmaker),专注于宏观 AI 战略以及亚洲地区的能力建设。我也是 Second Look Research 的联合创始人,并担任佐治亚理工 AI 安全倡议(Georgia Tech AI Safety Initiative)的顾问。

Previously, I

此前,我

Outside of work, I generally love being active. I used to play competitive golf and table tennis. I also like checking out new cafes and restaurants!

工作之外,我喜欢运动。我曾经打过竞技高尔夫和乒乓球。我也喜欢探店,尝试新的咖啡馆和餐厅!

Check out some of my amazing friends, collaborators, and mentors. Parv, Will, Zephy, Jasmine, Jo, Andrew, Andy, Rocio, Tzu, Mantas, Animesh, Kartik, Abdur Raheem, Max.

来认识一下我那些很棒的朋友、合作者和导师:ParvWillZephyJasmineJoAndrewAndyRocioTzuMantasAnimeshKartikAbdur RaheemMax

Research研究

I'm interested in understanding what and how general AI systems learn & generalize in order to develop scalable methods for alignment, evaluations, and control. My latest projects are to do with agentic misalignment in RL with CAIS, and interpretability of VLA models with the PAIR Lab. Some of my academic services include serving on the AAAI alignment track senior program committee, the Longitudinal Expert AI Panel (LEAP), and reviewing for workshops at ICLR, ICML, NeurIPS, and EMNLP. Research blogs and unpublished work are under writing.

我希望理解通用 AI 系统学到了什么、如何学习并泛化,从而开发可扩展的对齐、评估与控制方法。我最近的项目包括与 CAIS 合作研究强化学习中的智能体失调(agentic misalignment),以及与 PAIR Lab 合作研究 VLA 模型的可解释性。我的学术服务包括担任 AAAI 对齐方向的高级程序委员会委员、纵向 AI 专家小组(LEAP)成员,以及为 ICLR、ICML、NeurIPS 和 EMNLP 的研讨会审稿。研究博客和未发表的工作见写作页面。

A. Ghosh, A. Reuel, J. Chim, …, Y. Hao, …, I. Solaiman
arXiv preprint 2026arXiv 预印本 2026
Y. Hao, A. Panda, S. Shabalin, S. A. R. Ali
ICLR 2025 Human-AI Co-evolution Workshop + 2 others ICLR 2025 Human-AI Co-evolution 研讨会及另外 2 个
S. Shabalin, A. Panda, D. Kharlapenko, A. R. Ali, Y. Hao, A. Conmy
CVPR 2025 XIV WorkshopCVPR 2025 XIV 研讨会
G. Matlin, P. Mahajan, I. Song, Y. Hao, R. Bard, S. Topp, E. Montoya, M. R. Parwani, S. Shetty, M. Riedl
EMNLP 2025 Wordplay WorkshopEMNLP 2025 Wordplay 研讨会