Skip to content
Resume

Data Scientist / Software Engineer

Download PDF

I'm currently a data scientist with experience owning the full lifecycle of production decision systems, from feature and schema design through model-based and rules-engine scoring, ETL pipelines, and scalable serving. I enjoy collaborating with others to solve complex problems. Outside of work, I love developing fun projects, drawing, photography, hiking, and exploring food.

Data ScienceSoftware EngineeringMachine Learning

Experience

Quantifind

Palo Alto, CA

Data Scientist

Feb 2024 – Present
  • Designed and shipped a relationship-risk scoring system using a hybrid of model-based and rules-engine logic, balancing accuracy, interpretability, and operational constraints in production
  • Owned the full lifecycle of high-throughput data pipelines supporting sanctions and transaction screening, from design and methodology through production deployment and monitoring
  • Chose a pragmatic precompute-and-serve approach over a fully real-time model, shipping a pre-computed relationship-risk pipeline in Spark that enabled scalable, sub-second risk lookups in production
  • Led evaluation of unstructured address search methods and foreign-language risk model performance in ambiguous, not-well-defined problem spaces, translating open-ended questions into measurable approaches
  • Partnered across Platform, Product, and Front-End stakeholders to scope, prioritize, and ship relationship-network expansion initiatives, aligning technical and business requirements
  • Managed and mentored three interns, owning project scope, methodology, and delivery end to end

Associate Data Scientist

Sep 2022 – Aug 2024
  • Partnered directly with clients to design and operate production transaction-processing pipelines handling billions of records using Spark and Scala
  • Designed and built, from the ground up, the canonical schema and entity resolution logic for standardized structured relationship data used across downstream systems
  • Built end-to-end ETL pipelines with web scraping (Python, R) and Spark, including NER and location extraction, loading structured output into PostgreSQL
  • Optimized complex PostgreSQL queries and indexing strategies to improve relationship-retrieval performance at scale
  • Built a real-time knowledge graph service in Scala from zero to one, surfacing entity relationships on the fly while balancing latency and complexity trade-offs

Associate Data Scientist

May 2019 – Sep 2022
  • Built ETL pipelines, trade network visualizations, and tuned spaCy NER models

Education

University of Edinburgh

BSc Computer Science and Artificial Intelligence · Edinburgh, Scotland
2018 – 2022

Georgia Institute of Technology

MS Computer Science · Atlanta, Georgia
2025 – Present (Part-Time)

Skills

Languages

ScalaPythonSQLRJavaHaskell

Data & ML

NLPLLMspandasscikit-learnPyTorch

Engineering

GitDockerREST APIsAzkaban

Infra & Tools

AWS / GCPCI/CD