I’m a PhD researcher in LLM interpretability at Hessian.AI and the UKP Lab at TU Darmstadt. I’m interested in how language models work and what interpretability methods can actually establish.
No posts yet.