Chinmaya – Kubernetes, Ansible, Linux, experts in Lemon.io

Chinmaya

From Canada (UTC-5)flag

Site Reliability Engineer|Senior

Chinmaya – Kubernetes, Ansible, Linux

Chinmaya is a senior Site Reliability Engineer and Cloud Architect with deep expertise in Kubernetes, Azure AKS, Ansible, and large-scale cloud-native operations. He excels at managing multi-regional clusters, designing resilient architectures, and leading platform teams. His strengths include strategic infrastructure design, clear technical communication, and operational maturity in enterprise environments. Chinmaya is best suited for roles emphasizing platform reliability and scalability over niche storage or database administration.

15 years of commercial experience in
Aviation
Fintech
Retail
Telecommunications
Nonprofit
Main technologies
Kubernetes
15 years
Ansible
15 years
Linux
15 years
Python
15 years
Golang
15 years
Additional skills
Microsoft Azure
Apache Kafka
Docker
Knative
Jenkins
Prometheus
Grafana
Kafka
Redis
AI agent development
ElasticSearch
Terraform
Bash
GCP
Unix
GitLab CI/CD
Direct hire
Possible
Ready to get matched with vetted developers fast?
Let’s get started today!

Experience Highlights

Senior Site Reliability Engineer
Mar 2025 - Ongoing1 year 5 months
Project Overview

A global financial services organization providing investment banking, wealth management, securities, and investment management services to corporations, institutions, governments, and individuals worldwide.

Responsibilities:
  • Manage and support enterprise OpenShift platforms, including cluster builds, lifecycle management, upgrades, configuration, troubleshooting, and platform operations;
  • Lead GitOps-based application onboarding and deployment using FluxCD, establishing standardized and repeatable deployment patterns;
  • Manage enterprise Redis and Kafka platforms, including provisioning, configuration, security, onboarding, monitoring, troubleshooting, and lifecycle management;
  • Designed and implemented the observability stack, onboarding clients and applications with centralized metrics, logs, dashboards, and alerting;
  • Implemented and supported Kafka Connect integrations for reliable data movement between Kafka and external systems;
  • Conducted an end-to-end Elastic Cloud on Kubernetes (ECK) POC, validating key enterprise capabilities including mTLS, LDAP integration, scaling, Cross-Cluster Replication (CCR), security, and operational functionality;
  • Designed and implemented Istio Service Mesh, including Tier-2 gateways, East-West gateways, traffic management, and service-to-service connectivity for secure and scalable application communication;
  • Developed AI-assisted engineering capabilities, including skills and agents for automated Helm chart reviews, configuration validation, troubleshooting, and platform operations;
  • Automated repetitive platform operations and improved developer experience through GitOps, automation, and self-service capabilities;
  • Collaborate across application, DevOps, security, and infrastructure teams to establish platform standards, reliability practices, and scalable onboarding processes;
  • Drive platform modernization through Kubernetes/OpenShift, GitOps, observability, service mesh, automation, and AI-assisted operations.
Project Tech stack:
Kubernetes
Kafka
Redis
ElasticSearch
Prometheus
Grafana
Python
AI agent development
Senior Site Reliability Engineer
Feb 2022 - Feb 20253 years
Project Overview

A global technology provider delivering IT, communications, and digital solutions to the air transport industry, serving airlines, airports, and other aviation organizations worldwide.

Responsibilities:
  • Architected and managed enterprise-scale cloud and Kubernetes platforms across on-prem datacenters and Azure/AKS, supporting containerized applications and critical business services;
  • Led the design and evolution of Kubernetes platform architecture, including cluster provisioning, upgrades, networking, storage, security, observability, and operational tooling;
  • Built highly automated CI/CD and application onboarding platforms using Bamboo, Helm, Terraform, Ansible, and Kubernetes, significantly reducing manual deployment and operational effort;
  • Designed reusable, complex Helm-based deployment frameworks and established standardized patterns for application deployment, configuration, and lifecycle management;
  • Automated complex Kubernetes operations through custom Operators and Go-based automation, eliminating manual intervention and improving platform reliability;
  • Designed and implemented enterprise secret-management solutions using HashiCorp Vault and Azure Key Vault for secure application and infrastructure credential management;
  • Managed Azure infrastructure using Terraform, including AKS, networking, storage, databases, load balancers, Event Hubs, container registries, and hybrid connectivity through VPN and ExpressRoute;
  • Designed highly available Kubernetes storage solutions using NFS and Rook-Ceph to meet application availability and scalability requirements;
  • Established comprehensive observability and SRE practices using Prometheus, Alertmanager, New Relic, and Nagios, including monitoring, alerting, SLO/SLA management, and proactive incident detection;
  • Implemented Kubernetes security and governance through RBAC, authentication/authorization, Policy-as-Code, vulnerability scanning, and CI/CD security controls using SonarQube, Mend, and container scanning solutions;
  • Architected and managed service-to-service networking using Istio, including traffic management, security, and service observability;
  • Managed enterprise messaging platforms including Kafka, Confluent Kafka, and Azure Event Hubs, supporting reliable application integration and event-driven architectures;
  • Drove continuous platform modernization and Kubernetes upgrades, adopting new capabilities while maintaining security, availability, and operational stability;
  • Partnered with application, DevOps, security, and infrastructure teams to establish platform standards, engineering best practices, and scalable self-service capabilities.
Project Tech stack:
Apache Kafka
Microsoft Azure
Kubernetes
Python
Ansible
Terraform
DevOps Engineer
Jun 2021 - Mar 20229 months
Project Overview

A leading Canadian telecommunications and technology provider delivering wireless, broadband, TV, and digital services supported by large-scale technology and network infrastructure.

Responsibilities:
  • Developed, built, and managed Kubernetes container platforms (vanilla Kubernetes, OpenShift 4.6, OpenShift 4.8, Rancher) on bare-metal servers in private datacenters;
  • Performed platform builds and installations following the Infrastructure as Code approach using GitLab, Terraform, and Ansible;
  • Created end-to-end automation for tenant service deployments using GitOps, Helm charts, and Operators;
  • Architected platform security hardening and performed POCs on Prisma and Sysdig for container vulnerability scanning and reporting;
  • Created monitoring rules and service monitors for proactive alerting to the support team on issues and threshold breaches;
  • Performed POCs on public cloud Kubernetes platforms to identify suitable PaaS solutions for deploying Bell applications;
  • Worked on critical projects including 5G infrastructure builds and MEC (Multi-access Edge Computing);
  • Configured cloud-native storage solutions for containerized workloads through CSI using Trident and Rook-Ceph;
  • Configured Tigera Calico as the CNI for Kubernetes clusters and used its features for improved network management;
  • Integrated application deployment pipelines with SonarQube for vulnerability scanning.
Project Tech stack:
Kubernetes
GitLab CI
CD
Ansible
Platform Engineer
Apr 2020 - Jun 20211 year 1 month
Project Overview

A leading Canadian food and pharmacy retailer operating a broad portfolio of grocery, health, and consumer brands across the country.

Responsibilities:
  • Designed, built, and managed enterprise container platforms across GCP and Azure, leading the delivery of multiple Red Hat OpenShift and GKE clusters using automated and manual deployment approaches;
  • Led OpenShift platform engineering and lifecycle management, including cluster upgrades, node health, autoscaling, storage, networking, RBAC, authentication, authorization, and security;
  • Built enterprise observability and logging platforms using Prometheus, Grafana, Alertmanager, Elasticsearch, Fluentd, Kibana, Splunk, and AppDynamics, enabling proactive monitoring and automated incident management;
  • Developed Jenkins CI/CD pipelines and guided application teams on Kubernetes deployment strategies, S2I, and secure application exposure using OpenShift Routes;
  • Implemented platform security and compliance through SIEM, Qualys, OpenSCAP, network policies, RBAC, and secure cluster configurations;
  • Designed and tested GSLB-based disaster recovery and automated failover solutions for highly available business-critical applications;
  • Drove Infrastructure as Code and automation using Terraform, Python, Bash, and Shell scripting, automating cluster operations, backups, object cleanup, machine-set configuration, and health checks;
  • Managed Kubernetes/OpenShift storage architecture, including StorageClasses, PVs/PVCs, and application-specific storage requirements;
  • Established SLO/SLI-driven platform operations, cluster maintenance schedules, capacity management, cost tracking, and proactive issue resolution;
  • Led platform engineering teams using Agile methodologies, managing sprint backlogs, technical priorities, application escalations, and collaboration with Red Hat support/TAM on critical platform issues;
  • Maintained hands-on expertise with CoreOS, Podman, CRI-O, cgroups, namespaces, Custom Resources, Kubernetes Operators, and Service Mesh.
Project Tech stack:
Kubernetes
GCP
Microsoft Azure
Linux
Ansible
Unix
Bash
Senior System Support Analyst
Dec 2018 - Mar 20201 year 3 months
Project Overview

A regulatory organization responsible for governing lawyers and paralegals in Ontario, protecting the public interest through professional standards, regulation, and support for competence and ethical practice.

Responsibilities:
  • Planned, developed, implemented, deployed, and supported Red Hat Linux infrastructure and middleware application servers that were virtual or cloud-based;
  • Led and implemented new applications on Oracle WebLogic and Red Hat OpenShift platforms and provided post-implementation support;
  • Set up AppDynamics to monitor the OpenShift container platform;
  • Installed and updated OpenShift clusters on 3.x and 4.x versions and managed components such as EFK stack, Prometheus, Alertmanager, and Grafana;
  • Managed user authentication using RBAC, OAuth configuration, LDAP, and persistent storage such as PV and PVC in OpenShift and Kubernetes clusters;
  • Managed pods, routes, services, and SSL certificates in OpenShift and Kubernetes clusters;
  • Configured and managed Resource Quotas, Limit Ranges, services, routes, ConfigMaps, and Secrets on OpenShift and Kubernetes clusters;
  • Worked with advanced OpenShift concepts such as Serverless, Service Mesh, Knative, CoreOS, and Operators;
  • Set up Jenkins pipelines for build and deployment in OpenShift clusters;
  • Managed Docker images and created custom images using Dockerfiles;
  • Created automation scripts using Bash scripting and Ansible.
Project Tech stack:
Kubernetes
Jenkins
Docker
Ansible
Knative
Linux

Education

2011
BTECH
Bachelor's degree

Languages

English
Advanced

Hire Chinmaya or someone with similar qualifications in days
All developers are ready for interview and are are just waiting for your requestdream dev illustration
Copyright © 2026 lemon.io. All rights reserved.