Senior data engineer

Turning raw data
into refined, trusted
decisions.

I'm Mohammed Khan. I design the pipelines that carry a company's data from messy and raw to clean, governed, and ready to build on. Eight years in, most recently shipping the lakehouse behind The Pokémon Company's global analytics.

Databricks Certified Data Engineer Associate Ex-Pokémon, EX-Amazon
Portrait of Mohammed Khan
Databricks Certified · Data Engineer
Bronze layer, raw intake

The unprocessed facts, straight from the source.

Every good pipeline starts with an honest look at the raw data. Here's the unfiltered version of who I am.

I've spent the last eight years building the systems that turn a company's raw data into something people can actually trust and use. Most recently that meant leading platform engineering work at The Pokémon Company International, and before that, at Amazon Web Services, where I built and hardened the pipelines behind some of their internal analytics.

The work I enjoy most isn't glamorous. It's migrating tangled legacy ETL onto modern lakehouse platforms, building governance that people actually follow, and automating away the manual reconciliation that quietly eats up a team's week. I'm a Databricks Certified Data Engineer Associate, and I hold a Master's degree in Computer Science.

If you're looking for someone who can walk into a messy data environment and leave it dependable, I'd love to talk.

Outside of pipelines, I try to keep life just as well maintained. Gym, hiking, and cooking healthy meals are a big part of my week, I keep a couple of aquariums at home, and I'm a brand ambassador for Unmatched Supps.

CERTIFICATIONDatabricks Certified Data Engineer Associate
TOP SKILLS
Databricks Streamlining Complex Work Processes Apache Spark Python Scala AWS Apache Airflow SQL Server Alteryx
Silver layer, refined and conformed

Eight years of experience, cleaned up and cross-referenced.

The raw history, joined and structured into a career: four roles, three companies, and one throughline, moving data systems from fragile to dependable.

Senior Data EngineerIntellitask
NOV 2024 TO PRESENT · 1 YR 9 MOS · EDMONTON, AB
  • Engineered a high-fidelity migration strategy from a legacy CRM to Salesforce, holding near-perfect data accuracy while compressing the migration timeline.
  • Leading the client's move off their custom-built CRM platform onto Salesforce, mapping legacy objects and fields to Salesforce's data model so nothing gets lost along the way.
  • Building out data validation checks that catch quality issues before records ever land in Salesforce.
  • Architected Alteryx workflows and batch macros that automate ETL end to end, cutting manual intervention through server-side scheduling.
  • Optimized T-SQL stored procedures and query logic in SQL Server to speed up real-time data reconciliation.
  • Standardized cross-regional pipelines into scalable, reusable workflows, minimizing discrepancies and maintenance overhead.
Data Platform EngineerThe Pokémon Company International
APR 2021 TO NOV 2024 · 3 YRS 8 MOS · SEATTLE, WA
  • Spearheaded the enterprise migration of ETL workloads from AWS Glue to a Databricks Data Lakehouse, cutting operational cloud costs and lifting performance.
  • Built high-throughput pipelines with PySpark and Scala on AWS for faster, more reliable processing.
  • Implemented Databricks Unity Catalog to enforce data governance, mitigate compliance risk, and guarantee lineage.
  • Automated ingestion of complex XML and JSON datasets into Snowflake with custom Python, eliminating hundreds of hours of manual handling.
Senior Data EngineerAmazon
AUG 2019 TO MAR 2021 · 1 YR 8 MOS · SEATTLE, WA
  • Designed data pipelines and ELT processes on native AWS services (Redshift, S3, Glue) to improve processing reliability.
  • Led the migration from Redshift to a S3-based data lakehouse, enabling advanced analytics through QuickSight.
  • Wrote optimized Python and PySpark transformations within Glue and EMR to streamline complex workflows.
  • Moved orchestration from legacy internal tooling to Apache Airflow, and implemented encryption and IAM policy for PII with zero security breaches.
Big Data DeveloperGlobal Atlantic
JAN 2018 TO AUG 2019 · 1 YR 8 MOS · BOSTON, MA
  • Managed and optimized Hadoop ecosystems (Hortonworks, Cloudera) spanning Hive, Spark, and Kafka.
  • Migrated legacy Informatica ETL jobs onto Hadoop to lift processing capacity for large data volumes.
  • Built and tuned Spark Core RDD applications in Scala for large-scale data comparison and processing.
Gold layer, curated for impact

The version that goes to leadership.

Aggregated, business-ready, and easy to act on: the skills, credentials, and highlights that summarize eight years of pipeline work.

8+
YEARS IN DATA ENGINEERING
4
COMPANIES, ONE THROUGHLINE
3.7YR
LONGEST TENURE, POKÉMON CO.
0
PII SECURITY BREACHES ON WATCH

Core skills

Databricks Streamlining Complex Work Processes Apache Spark Python Scala AWS Glue / EMR / Redshift / S3 Apache Airflow SQL Server / T-SQL Alteryx Hadoop / Hive / Kafka

Certification

Databricks Certified Data Engineer Associate

Education

Governors State University
Master of Science, Computer Science
MAY 2016 TO DEC 2017
Jawaharlal Nehru Technological University
Bachelor of Computer Science

Let's build something that scales.

If you've got a data problem worth solving, I'd like to hear about it. Send me a message and I'll get back to you directly.