Piotr – Prometheus, Datadog, Kubernetes, experts in Lemon.io

Piotr

From Portugal (UTC+1)flag

Platform Engineer
Site Reliability Engineer

Piotr – Prometheus, Datadog, Kubernetes

Piotr is a senior Site Reliability and Platform Engineer with deep expertise in observability, Kubernetes operations, and reliability engineering. He has led implementations using Prometheus, Datadog, OpenTelemetry, and Terraform, and demonstrated strong troubleshooting and automation skills. His experience spans large-scale AWS environments and developer enablement.

18 years of commercial experience
Main technologies
Prometheus
1 year
Datadog
1.5 years
Kubernetes
6.5 years
AWS
5 years
CI/CD
4.5 years
Helm
1 year
Terraform
3.5 years
OpenTelemetry
1 year
Thanos
1 year
Python
1 year
Linux
1 year
Additional skills
Grafana
Amazon CloudFront
Big Data
Ansible
Puppet
DevOps
Network administration
System administration
FreeBSD
CloudWatch
ElasticSearch
Kibana
Direct hire
Possible
Ready to get matched with vetted developers fast?
Let’s get started today!

Experience Highlights

Senior Site Reliability Engineer (Observability)
Dec 2024 - Ongoing1 year 8 months
Project Overview

An all-in-one workplace platform combining project management, AI agents, and a shared company "brain" to help teams and software work together more efficiently, bringing all context and AI tools into one place.

Responsibilities:
  • Managed and optimized production workloads on EKS, with a strong focus on performance tuning and cost efficiency.
  • Built and maintained observability on DataDog, including dashboards, metrics, and logs from EKS workloads, CloudFront, and ALB.
  • Designed and implemented a self-hosted, end-to-end logging and observability stack powered by Loki and Grafana.
  • Developed scalable logging pipelines using AWS Glue, Athena, Vector, and DataDog for analytics and monitoring.
  • Delivered observability pipeline architecture to ensure reliable, real-time operational insights.
  • Administered and optimized AWS CloudFront CDNs to improve performance and edge delivery.
  • Participated in on-call rotations and led incident response in production environments.
Project Tech stack:
AWS
Kubernetes
Datadog
Grafana
Amazon CloudFront
ETL
AI-assisted coding
OpenTelemetry
Prometheus
Claude Code
Terraform
Terragrunt
Senior Site Reliability Engineer (Platform)
Oct 2023 - Nov 20241 year
Project Overview

A global online marketplace platform connecting buyers and sellers for second-hand goods, vehicles, real estate, and jobs, headquartered in Amsterdam.

Responsibilities:
  • Deployed and maintained EKS clusters at scale.
  • Implemented CI/CD pipelines with GitLab Pipelines for automated testing and deployment.
  • Utilized ArgoCD for GitOps practices, ensuring seamless continuous delivery.
  • Enhanced system reliability through automation.
  • Enforced security policies using Kyverno to ensure compliance standards.
  • Implemented monitoring solutions using NewRelic and Prometheus Stack.
  • Used Helm for packaging and deployment processes.
  • Used Kustomize for customized Kubernetes manifest management.
Project Tech stack:
Kubernetes
CI
CD
GitLab CI
CD
Prometheus
Helm
Apache Solr
Terragrunt
Terraform
AWS
Site Reliability Engineer
Mar 2020 - Oct 20233 years 7 months
Project Overview

A global online marketplace platform connecting buyers and sellers for second-hand goods, vehicles, real estate, and jobs, headquartered in Amsterdam.

Responsibilities:
  • Maintained the Apache Solr search platform.
  • Operated and maintained 40+ production Amazon EKS clusters across multiple AWS accounts and regions.
  • Kept infrastructure as code using Terragrunt and Terraform.
  • Kept sites operating reliably.
  • Used monitoring tools including Prometheus, Thanos, New Relic, SignalFX, and Sentry.
  • Used CI/CD tooling including GitLab Pipelines, Atlantis, and ArgoCD.
  • Worked with the EMR big data platform.
  • Designed and automated AWS infrastructure.
Project Tech stack:
Terraform
AWS
Kubernetes
Prometheus
CI
CD
Sentry
Big Data
Networking
GitLab CI
CD
Terragrunt
DevOps Consultant
Sep 2018 - Nov 20191 year 2 months
Project Overview

A company developing tools for real-time virtual reality streaming, enabling users to create and connect through interactive, immersive virtual reality experiences, including social streaming video and immersive 360° content.

Responsibilities:
  • Managed infrastructure as code using Terraform and Helm.
  • Maintained Kubernetes clusters on AWS using KOPS.
  • Monitored infrastructure using Prometheus (Alertmanager, Thanos) and Grafana.
  • Maintained MongoDB clusters, RabbitMQ clusters, S3-compatible object storages, and CDN storages.
Project Tech stack:
Kubernetes
Terraform
AWS

Education

2007
Information Technology
Maturity Exam

Languages

English
Advanced

Hire Piotr or someone with similar qualifications in days
All developers are ready for interview and are are just waiting for your requestdream dev illustration
Copyright © 2026 lemon.io. All rights reserved.