Gouki Minegishi

I’m a second-year PhD student at The University of Tokyo, mentored by Professor Yutaka Matsuo.

What I enjoy most is taking a model too complicated to reason about and finding a setting simple enough to explain completely — small models, controlled tasks, clean training regimes — then using what those settings reveal to work out how capabilities actually come about. The question underneath all of it is what human intelligence really is.

For more details, see my CV.

selected publications

  1. ICML2026
    analogy.png
    Emergent Analogical Reasoning in Transformers
    Gouki Minegishi, Jingyuan Feng, Hiroki Furuta, Takeshi Kojima, Yusuke Iwasawa, and 1 more author
    In Forty-third International Conference on Machine Learning, ICML, 2026
  2. ACL2026
    emergent_misalignment.png
    Understanding Emergent Misalignment via Feature Superposition Geometry
    Gouki Minegishi, Hiroki Furuta, Takeshi Kojima, Yusuke Iwasawa, and Yutaka Matsuo
    In The 64th Annual Meeting of the Association for Computational Linguistics, 2026
  3. Neurips2025
    Reasoning_Graph.gif
    Topology of Reasoning: Understanding Large Reasoning Models through Reasoning Graph Properties
    Gouki Minegishi, Hiroki Furuta, Takeshi Kojima, Yusuke Iwasawa, and Yutaka Matsuo
    In The Thirty-ninth Annual Conference on Neural Information Processing Systems, Neurips, 2025
  4. ICML2025
    ICL.gif
    Beyond Induction Heads: In-Context Meta Learning Induces Multi-Phase Circuit Emergence
    Gouki Minegishi, Hiroki Furuta, Shohei Taniguchi, Yusuke Iwasawa, and Yutaka Matsuo
    In Forty-second International Conference on Machine Learning, ICML, 2025
  5. ICLR2025
    SAE.png
    Rethinking Evaluation of Sparse Autoencoders through the Representation of Polysemous Words
    Gouki Minegishi, Hiroki Furuta, Yusuke Iwasawa, and Yutaka Matsuo
    In The Thirteenth International Conference on Learning Representations, ICLR, 2025