Kareem
From Egypt (UTC+2)
Kareem – Kubernetes, Python, Ansible
Kareem is a strong senior Site Reliability and DevOps Engineer with broad production experience and sound troubleshooting instincts. He demonstrates senior-level operational judgment, clear client-facing communication, and a strong ownership mindset. Kareem is experienced in working independently on complex production environments, identifying and resolving operational issues, and supporting reliable day-to-day platform operations. He is well-suited for senior DevOps/SRE roles requiring independent decision-making, effective collaboration with engineering teams, and a strong focus on platform reliability and operational stability.
19 years of commercial experience in
Main technologies
Additional skills
Direct hire
PossibleReady to get matched with vetted developers fast?
Let’s get started today!Experience Highlights
Senior SRE/ DevOps Engineer
A high-traffic e-commerce platform serving customers through a scalable web-based shopping application. The platform was undergoing a transition from a monolithic application architecture toward cloud-native microservices running on Kubernetes. The engineering environment included Azure Kubernetes Service, multi-cloud infrastructure, CI/CD automation, infrastructure as code, centralized observability, security automation, and internal developer platform capabilities.
- Contributed to Cross plane & Backstage - Self Developer Portal project.
- Automated CI/CD pipelines in Python, Bash scripts, and Terraform modules, increasing structure and deployment efficiency.
- Designed, implemented, and planned CI/CD pipelines on Azure DevOps, GitHub Actions, and Jenkins platforms.
- Defined deployment strategies for various edge use cases such as canary and blue/green strategies on ArgoCD.
- Implemented DevSecOps best practices including code linting and container security checks.
- Migrated FlaschenpostSE Webshop App service from a monolithic architecture to a microservices-based app on AKS.
- Managed and configured Spot.IO by migrating on-demand AKS nodes to spot-based nodes to optimize costs while maintaining performance.
- Developed and maintained cloud-based microservices on Kubernetes, Helm, and Kustomize, ensuring scalability and fault tolerance.
- Managed multi-cloud environments across AWS and Azure, leveraging Terraform and Vault for infrastructure as code and secrets management.
- Designed, created, and defined monitors, logs, and alert definitions for Datadog, Grafana, Prometheus, and ELK, and integrated PagerDuty for on-call management.
- Led incident management duty and root cause analysis, collaborating with teams to resolve issues efficiently.
- Defined caching strategies and implemented Cloudflare rules to optimize performance.
- Collaborated with development teams on application best practices for scalable, fault-tolerant, and secure application development.
- Documented infrastructure, migration projects, and microservices architectures to ensure knowledge sharing and maintainability.
- Collaborated with developers on code reviews and project strategy definition.
Linux System / DevOps Engineer
An autonomous driving research and development platform focused on supporting vehicle research, simulation, testing, and experimentation. The project required reliable infrastructure across cloud and on-premises environments to support distributed workloads, automated software delivery, containerized applications, and large-scale testing. The engineering environment combined Kubernetes, Docker, CI/CD automation, infrastructure as code, monitoring, networking, and virtualization technologies.
- Contributed to the Diginet autonomous driving project, developing and maintaining infrastructure to support autonomous vehicle research and testing.
- Designed, implemented, and maintained CI/CD pipelines using Jenkins, GitHub Actions, GitLab CI/CD, and Travis CI, automating deployment processes for cloud and on-prem environments.
- Developed infrastructure automation using Terraform, Ansible, and Kubernetes, streamlining provisioning and configuration management.
- Managed containerized applications with Docker, Kubernetes, Kustomize, Helm, and ArgoCD, ensuring scalability and reliability.
- Configured and maintained version control systems, implementing best practices for repository management and code collaboration.
- Designed and implemented monitoring and logging solutions with Prometheus, Icinga, Graphite, Nagios, CloudWatch, and Grafana, ensuring system health and performance optimization.
- Administered multi-cloud and virtualization environments, including AWS, GitLab, VMware vSphere, OpenStack, and Citrix, supporting both cloud-native and hybrid deployments.
- Managed and secured enterprise networks, configuring Cisco hardware and implementing networking protocols.
- Managed different operating systems and server administration across multiple Linux distributions and Windows Server, as well as Active Directory and Exchange services.
Senior System Engineer
Enterprise cloud and infrastructure environments supporting business applications and distributed services across cloud and on-premises infrastructure. The environment included virtualized infrastructure, AWS, Azure, Linux servers, enterprise databases, identity and access management, collaboration platforms, networking, security, and high-availability services. The role focused on maintaining reliable infrastructure and integrating multiple enterprise platforms and services.
- Designed and deployed Hyper-V-based infrastructure, provisioning and managing virtualized environments for enterprise applications.
- Administered cloud-native and on-premises distributed systems, including AWS, Linux servers, ERP, Oracle, Power BI, SQL BI, Azure, and Office 365, ensuring high availability and performance.
- Managed and configured enterprise IT solutions such as CRM, ADFS, Exchange, SharePoint, Group Policy, SCCM, WSUS, TFS, cluster load balancing, single sign-on, and DirSync, improving system integration and security.
- Administered Windows-based services including Active Directory Certificate Services, Active Directory Rights Management Services, Active Directory Federation Services, Active Directory Lightweight Directory Services, and IIS, optimizing identity and access management.
- Managed network hardware and security, configuring Cisco ASA firewalls, routers, switches, and FortiGate firewalls, ensuring robust network protection and performance.
IT Service Desk
A leading media and publishing group with operations across Saudi Arabia and international branches.IT infrastructure and support environment for a large media and publishing organization with operations across multiple locations. The environment supported business users and internal IT services, with a focus on maintaining system availability, resolving technical incidents, and providing reliable day-to-day IT operations.
- Provided first and second-level IT support using SCCM 2007, ensuring efficient troubleshooting and resolution of technical issues.
- Managed help desk incidents and service requests through Manage Engine Service Desk, maintaining high response and resolution times.
- Supported the IT infrastructure for a leading media and publishing group with operations across Saudi Arabia and international branches, ensuring system stability and business continuity.
System Engineer
Enterprise IT infrastructure supporting business applications, retail systems, and financial operations. The environment included physical and virtualized servers, Windows infrastructure, Active Directory, networking, databases, disaster recovery, and enterprise business applications. The infrastructure was designed to provide reliable and highly available services for business-critical workloads.
- Built and deployed Hyper-V-based infrastructure on IBM servers, ensuring high availability and scalability.
- Administered Windows-based services, including Active Directory, DNS, DHCP, IIS, TMG, Forefront Endpoint Protection, SQL Server, and Team Foundation Server to support enterprise operations.
- Managed VMware and Acronis Disaster Recovery solutions, ensuring business continuity and data protection.
- Maintained and supported NCR POS machines and Microsoft Dynamics AX & NAV 2009, optimizing retail and financial system performance.