I am a Member of Technical Staff at Thinking Machines Lab, working on multimodal pretraining science and visual agents.
Previously, I was a Research Scientist at Google DeepMind
working on Gemini multimodal and visual agents (image2code).
I got my Ph.D. in Computer Sciences from University of Wisconsin-Madison, advised by Prof. Yong Jae Lee.