I am a Ph.D. student at the Institute of Automation, Chinese Academy of Sciences, advised by Prof. Xu-Yao Zhang. I work on the mechanistic interpretability of large models and agents, treating internal-mechanism understanding as a first-class design lever rather than a post-hoc diagnostic. My central question is how the internal computations of foundation models give rise to reasoning---and when and why they fail---with the answers directly driving methods for reliable reasoning and reliable agents. I stress-test this programme in VLA and WAM embodied settings, aiming to move interpretable, reliable behaviour from paper to product.
Topics: Interpretable AI, Multimodal Reasoning, Self-Evolving Agents, Embodied Intelligence
Please feel free to reach out if you are interested in related topics.