ML engineer with 6+ years building LLM and retrieval systems in production across search, ranking, and evaluation.
Experience Senior ML Engineer — Acme AI 2022 — Present · San Francisco, CA• Built the model-serving layer powering 12M inferences/day at 99.9% uptime.
• Led retrieval-quality work that lifted answer accuracy 21% across 4 models.
• Shipped an offline eval harness (LLM-as-judge) adopted by 5 teams.
Machine Learning Engineer — DataForge 2019 — 2022 · Remote• Vector-search pipeline over 40M documents with sub-100ms retrieval.
• Cut model training cost 27% by moving feature pipelines to Spark.
Projects• OpenRAG — open-source retrieval-eval toolkit; 2.1k GitHub stars.
• LatencyLab — inference profiler behind the p95 latency win at Acme.
Education BSc Computer Science — State University 2014 — 2018 · Dean's List SkillsPyTorch, Transformers, RAG, LLM eval, vector search, Ray, Kubernetes, AWS
Standard
★★★★★★★★★★Our most popular reverse-chronological layout. Maximum recruiter readability, clean parsing in every major ATS.