Analytics Engineer · Seattle, WA

Sanket
Thakre

5+ years at Amazon building scalable data pipelines, dimensional models, and self-serve analytics platforms on AWS.

Data Ecosystem
Scroll
0
Years at Amazon
0
Projects Built
0
Technologies
0
Users Served
About

Background

Sanket Thakre
Sanket Thakre Analytics Engineer · Seattle, WA

I'm an Analytics Engineer with 5 years at Amazon in Seattle, designing and building scalable data pipelines, dimensional models, and self-serve analytics platforms used by business, finance, and ops teams.

My focus is the full AE stack, from ingestion and transformation (ETL/ELT, dbt, Airflow) to data modeling and the BI layer, with a strong emphasis on reliability, automation, and making data genuinely accessible.

I hold an MS in Computer Engineering from Cal State Fullerton and a BE in Electronics & Telecommunications from the University of Pune. I enjoy working at the intersection of engineering and business, building data systems that actually get used.

📍
Location
Seattle, WA
🏢
Current
BIE @ Amazon
🎓
Education
MS Computer Engineering, CSUF · BE E&T, Pune
✉️
Email
sanket.thakre3@gmail.com
Journey

The Road So Far

2013 - 2018
Grew up in India, studied engineering in Pune
Completed a BE in Electronics & Telecommunications at the University of Pune, where I first got into software and realized I liked building things that worked with data.
Aug 2019
Packed up and moved to the US
Landed in Fullerton, California to start my MS at Cal State Fullerton. New country, new city, new chapter.
2019 - 2021
Graduate school + first US work experience
Finished my MS in Computer Engineering while working as a Graduate Assistant. Got my first taste of professional work in the US supporting university research communications and operations.
Aug 2021
Joined Amazon and moved to Seattle
Got the call from Amazon and made the move to Seattle. Started as a Business Intelligence Engineer on the global fulfillment analytics team, thrown into the deep end from day one, and loved it.
2021 - 2025
Four years building at Amazon scale
Owned end-to-end data systems across global fulfillment, from pipelines that run daily FC operations to planning tools used by leadership across NA, EU, Japan, and India. Learned what it means to build for reliability and scale.
Jun 2025 - Present
New team: Amazon Pharmacy
Moved internally to Amazon Pharmacy, taking on a new domain with new challenges. Building analytics for pharmacy operations and kiosk launches across North America.
Now
Leveling up into Analytics Engineering
Five years in, I know what good data infrastructure looks like and what it takes to build it. Now expanding into the modern data stack: dbt, Airflow, data quality frameworks. Looking for the next challenge as an Analytics Engineer.
Stack

Technical Skills

Data Engineering
ETL / ELTdbtApache AirflowFivetranAirbyteDimensional ModelingStar SchemaSCD TypesData Quality
Languages
SQL (Advanced)PythonPandasNumPyPySparkMatplotlibBash
Cloud & Warehouses
Amazon RedshiftSnowflakeDatabricksAWS S3AWS GlueMySQLPostgreSQLOracle
BI & Visualization
Amazon QuickSightTableauPower BILookerJupyter Notebook
Tools & Practices
GitGitHub ActionsAgile / ScrumData ContractsPerformance TuningRoot Cause Analysis
Experience

Where I've Worked

Amazon
Aug 2021 - Present · Seattle, WA
Business Intelligence Engineer · Pharmacy Jun 2025 - Present
  • Built North America Amazon Pharmacy kiosk launch analytics, including replenishment and user analytics pipelines tracking daily orders, providing visibility into kiosk performance across launch markets.
  • Migrated and maintained contractual SLA data pipelines for pharmacy partners, preventing financial penalties by proactively detecting anomalies outside contract scope.
Business Intelligence Engineer · Global Fulfillment Aug 2021 – May 2025
  • Sole product owner of a real-time fulfillment risk and alerting platform that integrated 35 sites, rebuilt unstable ETL pipelines, and achieved an 84% reduction in data quality incidents; foundational pipelines later adopted org-wide by a separate fulfillment planning service.
  • Led modernization of a global rate-setting system supporting annual operational planning across NA, EU, Japan, and India; analyzed 9 interconnected stored procedures and migrated to centralized ETL pipelines with version control and automated triggers.
  • Architected a centralized KPI platform standardizing 95 metrics across fulfillment operations, powering 4 data products across 3 business units, reducing new analytics solution delivery from 6 months to 2 weeks, and cutting data accuracy incidents by 25%.
  • Inherited a business-critical operations metrics pipeline (Go codebase) after the maintaining engineer departed; resolved critical accuracy issues in 2 days and delivered BI solutions (dimensional models, ETL pipelines, QuickSight dashboards) across fulfillment operations teams.
Cal State Fullerton
Sep 2019 - May 2021 · Fullerton, CA
Graduate Assistant · Office of Research & Sponsored Projects
  • Collaborated with the Associate Vice President to design and produce the department's annual report, making it accessible to a campus community of 45,000+ students and faculty.
  • Maintained department website, managed communications and scheduling, and supported coordination of research workshops and events to promote campus-wide research initiatives.
Education

Academic Background

Cal State Fullerton
MS · Computer Engineering
2019 - 2021 · Fullerton, California
Database Systems Algorithms Software Engineering Graduate Assistant
University of Pune
BE · Electronics & Telecommunications
2013 - 2018 · Pune, India
Digital Systems Signal Processing Software Development Pune, India
Projects

What I've Built

Python dbt Airflow
01
End-to-End ELT Pipeline with dbt & Airflow
PythondbtAirflowPostgreSQLGitHub Actions
9.5M rows of NYC TLC trip data ingested into PostgreSQL, transformed with dbt (staging → intermediate → dimensional mart), orchestrated with Apache Airflow, validated with 22 passing dbt tests, and monitored with Slack alerts. Fully containerized with Docker.
View on GitHub →
Raw Table PostgreSQL GE Checks Null · Schema · Stats Alert ✓ Pass
02
Automated Data Quality & Monitoring Framework
PythonpytestPandasSQL
Custom Python framework for automated data quality monitoring: schema validation, null/duplicate detection, referential integrity checks, and statistical outlier alerting. Generates HTML and JSON reports. 6/6 pytest tests passing, CI via GitHub Actions.
View on GitHub →
Revenue Churn Cohort
03
Self-Serve Analytics Platform for E-commerce
SQLdbtRedshiftQuickSight
Raw e-commerce data → dbt dimensional model → governed metric layer → interactive dashboard (revenue, churn, cohort retention) with documented data lineage.
View on GitHub →
Contact

Let's work together.

I build reliable data pipelines, dimensional models, and analytics platforms that help teams make better decisions. If you're working on something interesting in the data space, grab a time on my calendar and let's talk.