I am a research scientist at Google and a member of the Gemini large-scale pretraining core team. My work focuses on making language model evaluation more reliable: identifying and mitigating data contamination, improving evaluation signal-to-noise ratio, and understanding how training stages affect memorization and generalization. I also work on efficient, task-specific language models trained with synthetic data.
biography
I completed my PhD in Computer Science at Boston University in 2025 under the supervision of Prof. Derry Wijaya. My doctoral research covered LLM evaluation, machine translation, sentence representations, compositional generalization, and computational social science.
Before joining Google, I held research internships at Google, FAIR at Meta, and Amazon. I received my BSc in Electrical and Electronics Engineering from Bogazici University.