Data Engineer, 6+ Years

Data, in whatever shape it comes in.

I build resilient ETL/ELT pipelines, APIs, and databases. Before data engineering, I worked directly with environmental data, in the field and in the lab. That's a big part of why I'm comfortable with data that doesn't arrive in clean rows.

Julia Burmistrova
Python SQL Kafka Snowflake Airflow Spark dbt Databricks AWS Terraform Docker PostgreSQL CI/CD

Projects

Spark + dbt · Databricks · Prediction Markets 2026

World Cup prediction market pipeline

I built this as two parallel pipelines over Manifold Markets' public API and Polymarket, testing calibration and the favorite-longshot bias in 2026 World Cup markets. One runs Spark, dbt, and DuckDB or Postgres as a Kubernetes batch job; the other runs the same logic on Databricks with Delta Live Tables and Unity Catalog.

View project →
SAR + Optical · Remote Sensing 2022

Sensor fusion for inundation classification

I combined SAR and optical satellite imagery to classify flood inundation across California wetlands.

View project →
Field + Lab · Waste Systems 2019

Food waste & wastewater co-digestion

I studied co-digestion of food waste and wastewater solids in Yosemite National Park.

View project →

About. I'm a data engineer with 6+ years of experience. I build ETLs for third-party data, APIs, and databases, and I care about resiliency and efficiency more than cleverness.

I like working across varied, messy data sources. Before this, I worked with geospatial data for ecosystem monitoring: collecting samples in the field and running analysis in the lab.

These days I mostly prefer to stay indoors.

Get in touch

Open to conversations about data engineering, environmental monitoring, or anything in between.

Programming / Coding Laboratory Analysis Fieldwork Remote Sensing Teaching & Mentorship