Ryota Takatsuki
高槻 瞭大
I am a Ph.D. student at the Sussex Centre for Consciousness Science, University of Sussex, supervised by Anil Seth and Christopher Buckley, and a research fellow at the AI Alignment Network in Tokyo.
I study transformer-based language, vision, and vision–language models as comparative experimental systems for consciousness science. My work asks whether these models exhibit consciousness-associated phenomena, such as bistable perception and metacognition, and uses mechanistic interpretability to characterize how they are implemented. I see this as one tractable route into which sentience-relevant functions AI systems may already approximate, and which they do not.
Before Sussex, I was a research intern in Ryota Kanai’s research team at Araya Inc., where I worked on meta-representations as a way of operationalizing higher-order theories of consciousness for deep learning models, alongside the mechanistic interpretability of visual illusions. I also studied bistable perception in multimodal models at the R&D Center for Large Language Models, National Institute of Informatics, and real-time 3D video transmission at the University of Tokyo, where I completed my B.S. in Systems Innovation.
News
- Aug 2026
- Two papers accepted at EMNLP 2026: “(How) Do MLLMs Report Bistable Images Like Humans?” (main conference, first author; code) and “In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?” (Findings; code).
- Jul 2026
- Our paper on in-context neurofeedback was selected as a spotlight at the Trustworthy AI for Good workshop at ICML 2026 in Seoul (code).
- Jul 2026
- Presented a poster, “Do Autoregressive LLMs Exhibit Postdiction-Like Integration of Later Tokens?”, at ASSC 29 in Santiago, Chile, supported by an ASSC 29 Travel Award.
- 2026
- Our survey “Mechanistic Interpretability: A New Trend in Interpretability Research” (in Japanese, with Koshiro Aoki, Gouki Minegishi, and Hyunsoo Cho) received the JSAI 40th Anniversary Commemorative Paper Award, Best Paper.
- Oct 2025
- “Meta-representations as representations of processes” (with Ryota Kanai and Ippei Fujisawa) was published in Neuroscience of Consciousness.
- Oct 2025
- Our Academist Prize (5th cohort, 2025) crowdfunding project with Shosuke Nishimoto, “Understanding the World as Perceived by AI”, closed with 238 supporters and 192% of its goal.
- Sep 2025
- Started my Ph.D. at the Sussex Centre for Consciousness Science, University of Sussex, supervised by Anil Seth and Christopher Buckley.
- Jun 2025
- “Decoding Vision Transformers: the Diffusion Steering Lens” appeared at the First Workshop on Mechanistic Interpretability for Vision at CVPR 2025.
- May 2025
- Gave a talk on reverse-engineering the Kanizsa illusion in Vision Transformers at JSAI 2025 in Osaka, where I also served on the organizing committee of the organized session on mechanistic interpretability.
- Jun 2024
- Joined the AI Alignment Network as a research fellow.
Selected publications
- Takatsuki, R., Doi, T., Watahiki, A., Seth, A. K., & Yanaka, H. (2026). (How) Do MLLMs Report Bistable Images Like Humans? EMNLP 2026, main conference (to appear). Code
- Kanai, R., Takatsuki, R., & Fujisawa, I. (2025). Meta-representations as representations of processes. Neuroscience of Consciousness, 2025(1), niaf038. DOI
- Takatsuki, R., Joseph, S., Fujisawa, I., & Kanai, R. (2025). Decoding Vision Transformers: the Diffusion Steering Lens. Workshop on Mechanistic Interpretability for Vision, CVPR 2025. arXiv
- Aoki, K., Takatsuki, R., Minegishi, G., Haruki, Y., & Kawahara, D. (2026). In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access? Findings of EMNLP 2026 (to appear). Code