hi, i'm sagnik
i'm passionate about AI safety and post-training alignment of large language models.
my current work focuses heavily on mechanistic interpretability i.e understanding the internal circuits that drive model behavior.
things which i am researching / building stuff , I write over here :
blog · projects


