About Me

I'm currently building across AI safety evaluation in LLMs and physical AIs, mechanistic interpretability, and AI4Astro. My research interest lies in understanding how models work internally (layer-wise) and how this information can advance our progress in AI, while making sure they're safe to deploy in real-world settings. And most of the time, I'm very obsessive and competitive about learning and solving things that I really enjoy (especially in research).

Places I've worked in: EleutherAI, UniverseTBD, Shanghai Jiao Tong Uni., Uni. of Maryland, Metacreation Lab, SEACrowd.

Working Papers / Projects

Publications (preprints/peer-reviewed papers)

Projects