NLP interpretability with Dr. Divya Chaudhary's group; joint paper at TrustNLP, NAACL 2026.
Research
I study the failure modes that surface only when systems meet the real world: reward hacking under distribution shift, sycophancy and specification gaming, silent failures in latent reasoning, and the institutional mechanisms that catch what technical safeguards miss.
It has produced first-author workshop papers at ICLR, NeurIPS, AAAI, ICML and NAACL, including an oral presentation at the ICLR 2026 AI for Peace Workshop and joint work with the InsightX Lab at TrustNLP, and contributions to two EvalEval Coalition studies published at the ICML 2026 main conference. The complete, current record lives in one place: Google Scholar. Longer-form essays and research notes are on Substack.
Positions & Fellowships
Mechanistic interpretability as a lens for understanding model deviation from sanctioned behaviour.
Integrity disclosures for generative AI in democratic information environments.
Alignment research on RL agents under the mentorship of Prof. Fernando Rosas (University of Sussex), working towards ICLR 2027.
Led a project on prompt optimisation for verifiable hallucination reduction.
Designing synthetic datasets for RL-style post-training and evaluation under controlled task distributions.
Governance research for the Berkeley AI Safety Initiative.
Taught and supervised AI, machine learning, deep learning and systems courses and labs for B.Tech and dual-degree students; developed curriculum and assessment materials; organised the AMRIT-2023 and MINDS-2023 conferences.
Summer internships: sentiment-analysis pipeline at LIT, Bhubaneswar (2019); database indexing systems at Naresh i Technologies, Hyderabad (2018).
Education
CGPA 9.38/10 · Gold Medalist, Summa Cum Laude · Batch topper.
CGPA 8.68/10 · Ranked in the top 5 of the college.
Hackathons & Competitions
Project later accepted as a NeurIPS 2025 workshop paper.
Research Grants
Mechanistic interpretability: experimental design, model analysis and dissemination.
Pilot experiments and preliminary analyses for an independent research contribution.
Recognition
Accepted on the strength of a mechanistic-interpretability hackathon project. Continuing collaboration with Amirali Abdullah (Martian).
Selected with a partial scholarship; withdrew due to funding constraints.
Media
Coverage in Bloomberg Opinion.
Service & Volunteering
Workshop reviewing at ICLR, ICML and ACL, and conference reviewing for COLM.
Selected Programmes
Beyond Research
Languages: English (full professional), Odia (native), Hindi (limited working), Sanskrit (limited working), Spanish (elementary).
Off hours: critically acclaimed podcasts.
Work With Me
I'm open to research contractor roles, PhD opportunities, and collaborations in alignment, evaluations and AI governance. If your work touches any of the problems above, the fastest route is email.