about
👋 Hello, I am an Applied Research Scientist on the AMD GenAI team, where I work on multimodal foundation models. My research spans multimodal large language models (VLMs / MLLMs), data-centric AI, multimodal reasoning, and efficient learning.
At AMD, I work hands-on across the full foundation-model stack — large-scale data curation and data engines, pretraining, supervised fine-tuning, multimodal reasoning, and rigorous evaluation — with distributed training on AMD Instinct GPUs. I am a main contributor to InstellaVL-1B, the first fully-open multimodal LLM trained end-to-end on AMD GPUs, released with open weights, code, and a technical blog. I also contribute across the broader AMD Instella family of open models, including the Instella open language models and the Instella-T2I text-to-image model, spanning multimodal understanding, generation, and reasoning.
I earned my Ph.D. in Computer Science from Boston University (2019–2024), advised by Prof. Kate Saenko, with research internships at Meta AI, Google Cloud, and IBM Research. Earlier, I received my M.S. from the University of Michigan, Ann Arbor and my B.Eng. from Beijing University of Posts and Telecommunications.
