Liquid AI — ML Engineer Intern, GPU Inference
May 2026 — present · San Francisco, CA
- Bring new model architectures into SGLang, vLLM, TensorRT-LLM, llama.cpp, and Transformers, then profile and optimize their serving paths.
- Work across model architecture, CUDA graphs, cache design, kernels, constrained decoding, and low-precision numerics.
Shopify — ML Engineer Intern, Search
Sep — Dec 2025 · Toronto, ON
- Built and A/B tested a typo-correction pipeline for Shop.app search.
- Improved query rewriting with supervised fine-tuning and reinforcement learning against the strongest existing baseline.
- Ran large-scale retrieval and relevance evaluations across head queries.
Yupp AI — Software Engineer Intern
Feb — May 2025 · Mountain View, CA
- Built an LLM tool-calling framework with a modular registry, external retrieval connectors, and Google API integrations.
- Built the backend for pairwise preference collection, spanning routing, judge evaluation, labeling, quality control, and enrichment.