Sohan Venkatesh

preview.png

Hey there! I’m Sohan. I’m currently a research fellow at LASR Labs and am based in London.

I mainly work on AI safety and interpretability with the goal of reducing catastrophic risks caused by AI systems. To that end, I have previously worked on various projects including my recent papers on steering vectors and on mechanistic interpretability. You can find more about my research work here. Feel free to email me if you find my work interesting!

I’m motivated by effective altruism and I see AI safety as the most altruistic use of my technical skills. Outside of work/research, you will often find me reading, ranting about AI risks or going down Wikipedia rabbit holes at 2am (try it — follow the first link on any Wikipedia article 20 times and you’ll always end up on ‘Philosophy’).

news

20 Jul 2026 Started as a research fellow at LASR Labs, London.
11 Jun 2026 My papers on circuit localization and emotional valence in LLMs have been accepted at ICML Mechanistic Interpretability Workshop 2026!
21 Apr 2026 I will be heading to Rio to present my paper at ICLR Re-Align and CAO workshops!
15 Apr 2026 Awarded a $700 grant from BlueDot Impact to work on Chain-of-Thought faithfulness.
23 Mar 2026 Completed BlueDot Impact Technical AI Safety Course!

selected publications

Re-Align@ICLR

On the Non-Identifiability of Steering Vectors in Large Language Models

Sohan Venkatesh, Ashish Mahendran Kurapath

Accepted at the Representational Alignment Workshop, ICLR 2026

PDF
MechInterp

Architecture, Not Scale: Circuit Localization in Large Language Models

Sohan Venkatesh

Accepted at the Mechanistic Interpretability Workshop, ICML 2026

PDF
MechInterp

Negative Before Positive: Asymmetric Valence Processing in Large Language Models

Sohan Venkatesh

Accepted at the Mechanistic Interpretability Workshop, ICML 2026

PDF

latest posts