Resume
Download PDF
Data Scientist / Software Engineer
I'm currently a data scientist with experience owning the full lifecycle of production decision systems, from feature and schema design through model-based and rules-engine scoring, ETL pipelines, and scalable serving. I enjoy collaborating with others to solve complex problems. Outside of work, I love developing fun projects, drawing, photography, hiking, and exploring food.
Data ScienceSoftware EngineeringMachine Learning
Experience
Quantifind
Palo Alto, CA
Data Scientist
Feb 2024 – Present- Designed and shipped a relationship-risk scoring system using a hybrid of model-based and rules-engine logic, balancing accuracy, interpretability, and operational constraints in production
- Owned the full lifecycle of high-throughput data pipelines supporting sanctions and transaction screening, from design and methodology through production deployment and monitoring
- Chose a pragmatic precompute-and-serve approach over a fully real-time model, shipping a pre-computed relationship-risk pipeline in Spark that enabled scalable, sub-second risk lookups in production
- Led evaluation of unstructured address search methods and foreign-language risk model performance in ambiguous, not-well-defined problem spaces, translating open-ended questions into measurable approaches
- Partnered across Platform, Product, and Front-End stakeholders to scope, prioritize, and ship relationship-network expansion initiatives, aligning technical and business requirements
- Managed and mentored three interns, owning project scope, methodology, and delivery end to end
Associate Data Scientist
Sep 2022 – Aug 2024- Partnered directly with clients to design and operate production transaction-processing pipelines handling billions of records using Spark and Scala
- Designed and built, from the ground up, the canonical schema and entity resolution logic for standardized structured relationship data used across downstream systems
- Built end-to-end ETL pipelines with web scraping (Python, R) and Spark, including NER and location extraction, loading structured output into PostgreSQL
- Optimized complex PostgreSQL queries and indexing strategies to improve relationship-retrieval performance at scale
- Built a real-time knowledge graph service in Scala from zero to one, surfacing entity relationships on the fly while balancing latency and complexity trade-offs
Associate Data Scientist
May 2019 – Sep 2022- Built ETL pipelines, trade network visualizations, and tuned spaCy NER models
Education
University of Edinburgh
BSc Computer Science and Artificial Intelligence · Edinburgh, Scotland
Georgia Institute of Technology
MS Computer Science · Atlanta, Georgia
Skills
Languages
ScalaPythonSQLRJavaHaskell
Data & ML
NLPLLMspandasscikit-learnPyTorch
Engineering
GitDockerREST APIsAzkaban
Infra & Tools
AWS / GCPCI/CD