Hi, I'm Yugandhar ✌️ ! I'm a CS grad from the University of Southern California, and a Research Member and Collaborator at Elemental Research Lab. I keep coming back to one question: when a model says something, does it actually know it, or is it just producing the right shape of an answer? That question is what pulled me into working across mechanistic interpretability, world models, agent foundations, and mixture-of-experts architectures. The common thread is simple: understand what's actually happening inside a system well enough to reason about whether it's safe, not just whether it behaves. That's the standard I hold my own work to, and the one I think AI safety research needs more of.
I did my Bachelor's in CS at SRM, with research supervised by Prof. Alice Nithya. I interned at the University of St Andrews, and was a Research Contributor at Trinity College Dublin. Previously, I was a Software Engineer Intern at HydroMind and interned at IETE.
A test for whether a model's answer matches what it actually represents internally, or just sounds right.
Read Knowing vs. Saying →For full publication list, please refer to my Google Scholar page.