Zafer – Apache Spark, SQL, Python, experts in Lemon.io

Zafer

From Ireland (UTC+1)flag

Data Engineer|Senior

Zafer – Apache Spark, SQL, Python

Zafer is a Senior Data Engineer with approximately 14 years of experience in data platform architecture, large-scale ETL, and AWS-based solutions. He has led the design and migration of high-throughput data pipelines, including Spark-to-Kubernetes orchestration and GDPR compliance projects. Feedback highlights strong system-level thinking, practical delivery, and effective cross-team communication.

15 years of commercial experience in
Banking
Consumer goods
Cybersecurity
Data analytics
E-commerce
Insurance
Marketing
Software development
Main technologies
Apache Spark
6 years
SQL
14 years
Python
6 years
Airflow
3 years
AWS
6 years
ETL
13.5 years
Terraform
4.5 years
Additional skills
Snowflake
CI/CD
Kubernetes
Docker
PySpark
Business intelligence
Data Warehouse
Apache Kafka
DBT
Java
Scala
Data Modeling
Oracle
PL/SQL
AI-assisted coding
Amazon S3
API Gateway
JWT
Apache Hadoop
REST API
Hive
Bash
Database Management Systems
Direct hire
Possible
Ready to get matched with vetted developers fast?
Let’s get started today!

Experience Highlights

Senior Data Engineer
Apr 2023 - Nov 20252 years 7 months
Project Overview

A high-throughput cybersecurity analytics platform processing approximately 50TB of data daily. It provides near real-time analytics capabilities for security operations. I led the data engineering architecture, owning the complete migration from legacy batch processing to a low-latency streaming infrastructure.

Responsibilities:
  • Orchestrated the shift from a legacy batch ETL system to a high-throughput near real-time processing infrastructure.
  • Achieved and maintained a 15-minute latency SLO for about 50TB/day of cybersecurity analytics data.
  • Spearheaded technical coordination with internal consumers and dependent parties.
  • Implemented platform engineering standards using Infrastructure-as-Code with Terraform.
  • Designed container orchestration workflows with Docker and Kubernetes.
Project Tech stack:
AWS
Terraform
Docker
Kubernetes
CI
CD
Apache Spark
SQL
Python
ETL
Senior Data Engineer
Apr 2021 - Apr 20232 years
Project Overview

A scalable data discovery and lakehouse search platform providing high-performance queryable data services for internal consumers. I acted as the lead engineer and domain expert, owning the infrastructure architecture from the S3 data layer to the containerized deployment workflows.

Responsibilities:
  • Engineered a scalable data lake search platform utilizing Trino over AWS S3.
  • Spearheaded technical coordination with internal consumers and dependent parties.
  • Implemented platform engineering standards using Infrastructure-as-Code with Terraform.
Project Tech stack:
Python
SQL
AWS
Terraform
CI
CD
API Gateway
JWT
Amazon S3
Data Engineer
Jul 2020 - Apr 20218 months
Project Overview

An enterprise data modernization initiative migrating on-premise legacy systems to a scalable cloud architecture. It decouples and transfers datasets to resolve structural scalability limits.

Responsibilities:
  • Executed a complete lift-and-shift of an on-premise legacy ETL system (Hadoop, Teradata, PL/SQL, in-house orchestration) to the cloud.
  • Engineered the migration of large-scale data pipelines from legacy Teradata-backed Hadoop clusters.
  • Operated PySpark workloads directly on YARN.
  • Orchestrated the decoupling and transfer of on-premise datasets to AWS S3 and Snowflake using Apache Airflow.
Project Tech stack:
Apache Spark
PySpark
AWS
Airflow
Snowflake
ETL
Python
Apache Hadoop
Software Engineer
Jul 2019 - Jul 20201 year
Project Overview

An enterprise-wide metadata management platform that establishes end-to-end data lineage across the organization's data assets. It serves internal data teams to track data provenance and discovery.

Responsibilities:
  • Developed an in-house data catalog project using Java from a forked Apache Atlas repository.
  • Deployed and configured the Apache Atlas infrastructure.
  • Implemented functionalities to manage data dictionaries, flag PII, and categorize sensitive data.
Project Tech stack:
Java
AWS
Terraform
REST API
Data Engineer
Aug 2018 - Jul 201911 months
Project Overview

A data privacy and regulatory compliance project aimed at identifying, anonymizing, and removing Personally Identifiable Information (PII) from legacy platforms.

Responsibilities:
  • Contributed to GDPR compliance efforts by migrating and anonymizing sensitive user records within the Hadoop ecosystem.
  • Applied GDPR practices for the anonymization of legacy data according to C1/C2 data privacy classifications.
  • Designed specific data models for handling and structuring sensitive data securely.
Project Tech stack:
SQL
Apache Hadoop
Hive
Database Management Systems
Bash
Lead BI / DWH Consultant
Jan 2014 - Aug 20184 years 6 months
Project Overview

Enterprise data warehouse and business intelligence implementations for banking and insurance clients.

Responsibilities:
  • Directed end-to-end Enterprise Data Warehouse and Business Intelligence implementations for banking and insurance clients.
  • Architected Data Marts in adherence to Inmon and Kimball dimensional modeling methodologies.
  • Optimized mission-critical EDW refresh processes through query tuning and ETL redesign, reducing nightly load durations from 13 hours to 7 hours.
  • Led the technical and communication parts of the projects to ensure that deliveries are valid and on track.
Project Tech stack:
Data Warehouse
Business intelligence
ETL
SQL

Education

2009
Computer Engineering
B.Sc.

Languages

Turkish
Intermediate
English
Advanced

Hire Zafer or someone with similar qualifications in days
All developers are ready for interview and are are just waiting for your requestdream dev illustration
Copyright © 2026 lemon.io. All rights reserved.