NanoBanana侧面黑白鲨守-modified.png

🛬 Attending

🇰🇷ICML2026, Seoul, KR


My research is supported by:

Coefficient Giving

TML

Zulip


<aside> <img src="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-806f-bf66-007a2758f01d" alt="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-806f-bf66-007a2758f01d" width="40px" /> Google Scholar

</aside>

<aside> <img src="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-8076-afbe-007af223fbba" alt="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-8076-afbe-007af223fbba" width="40px" /> Research Gate

</aside>

<aside> <img src="attachment:5a1e77b6-a69a-4a9c-8282-b0e70b4de887:Less_Wrong_LOGO.png" alt="attachment:5a1e77b6-a69a-4a9c-8282-b0e70b4de887:Less_Wrong_LOGO.png" width="40px" /> LessWrong

</aside>

Socials


<aside> <img src="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-8012-8d55-007a1a6ff476" alt="notion://custom_emoji/0d39d0ab-438c-4f29-be70-03aa9d912057/2c928fcd-40c2-8012-8d55-007a1a6ff476" width="40px" />

Linkedin

</aside>

I work on AI trust for (multi-) agentic system that communicates, reasons and acts autonomously, making such system smarter (reasoning) and safer (alignment) at the same time. My reasoning work primarily focus on multimodal physics-physical reasoning with the applications in AI for Physical Science and Embodied Intelligence. My safety work revolves around spontaneous misalignment (such as sycophancy, deception and collusion).

My vertical focus on reliability and generalizability of evaluation and post-training, where I often leverage horizontal methods from interpretability and formal verification in Lean:

Useful stuff I (personally) recommend: