About Me

I am a fourth-year Ph.D. candidate at Dartmouth College. My research focuses on machine learning problems at the intersection of computer vision, large language models, multimodal reasoning, biomedical AI, and clinical natural language processing, with a current focus on multimodal learning for video, audio, and language understanding.

I am currently interning at Amazon logo Amazon on multimodal vision-language models and diffusion models with the GenAI Evaluation Media (GEM) team (Summer 2026), and I interned at Genentech logo Genentech on diffusion and flow matching models for omics data with the Biology Research AI Development (BRAID) lab (Fall 2025).

My recent publications span multimodal understanding, audio-language reasoning, digital pathology, medical knowledge reliability in large language models, and digital health applications. You can find the full publication list on my Google Scholar profile .

Teaching & Service

  • Teaching: Graduate TA for Machine Learning and Algorithms for Statistical Learning for Big Data.
  • Reviewers: CVPR (🏆 Outstanding Reviewer), ICML (🏅 Gold Reviewer), ECCV, ICCV, ACL, EMNLP, NeurIPS, ICLR, ACM MM, IJCNN, MICCAI, and ICASSP.

Research Interests

  • Multimodal models and reasoning
  • Computer vision and visual-language learning
  • Biomedical AI and digital pathology
  • Clinical NLP and digital health
  • Reliable and medically grounded LLMs

News

Publications

Multimodal Models

Biomedical AI and Digital Pathology

Clinical NLP and Digital Health

Medical LLMs and Evaluation