Ryota Takatsuki

高槻 瞭大

I am a Ph.D. student at the Sussex Centre for Consciousness Science, University of Sussex, supervised by Anil Seth and Christopher Buckley, and a research fellow at the AI Alignment Network in Tokyo.

Ryota Takatsuki

I study transformer-based language, vision, and vision–language models as comparative experimental systems for consciousness science. My work asks whether these models exhibit consciousness-associated phenomena, such as bistable perception and metacognition, and uses mechanistic interpretability to characterize how they are implemented. I see this as one tractable route into which sentience-relevant functions AI systems may already approximate, and which they do not.

Before Sussex, I was a research intern in Ryota Kanai’s research team at Araya Inc., where I worked on meta-representations as a way of operationalizing higher-order theories of consciousness for deep learning models, alongside the mechanistic interpretability of visual illusions. I also studied bistable perception in multimodal models at the R&D Center for Large Language Models, National Institute of Informatics, and real-time 3D video transmission at the University of Tokyo, where I completed my B.S. in Systems Innovation.

News

Aug 2026
Two papers accepted at EMNLP 2026: “(How) Do MLLMs Report Bistable Images Like Humans?” (main conference, first author; code) and “In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?” (Findings; code).
Jul 2026
Our paper on in-context neurofeedback was selected as a spotlight at the Trustworthy AI for Good workshop at ICML 2026 in Seoul (code).
Jul 2026
Presented a poster, “Do Autoregressive LLMs Exhibit Postdiction-Like Integration of Later Tokens?”, at ASSC 29 in Santiago, Chile, supported by an ASSC 29 Travel Award.
2026
Our survey “Mechanistic Interpretability: A New Trend in Interpretability Research” (in Japanese, with Koshiro Aoki, Gouki Minegishi, and Hyunsoo Cho) received the JSAI 40th Anniversary Commemorative Paper Award, Best Paper.
Oct 2025
Meta-representations as representations of processes” (with Ryota Kanai and Ippei Fujisawa) was published in Neuroscience of Consciousness.
Oct 2025
Our Academist Prize (5th cohort, 2025) crowdfunding project with Shosuke Nishimoto, “Understanding the World as Perceived by AI”, closed with 238 supporters and 192% of its goal.
Sep 2025
Started my Ph.D. at the Sussex Centre for Consciousness Science, University of Sussex, supervised by Anil Seth and Christopher Buckley.
Jun 2025
Decoding Vision Transformers: the Diffusion Steering Lens” appeared at the First Workshop on Mechanistic Interpretability for Vision at CVPR 2025.
May 2025
Gave a talk on reverse-engineering the Kanizsa illusion in Vision Transformers at JSAI 2025 in Osaka, where I also served on the organizing committee of the organized session on mechanistic interpretability.
Jun 2024
Joined the AI Alignment Network as a research fellow.

Selected publications

All publications and presentations →