

Senior Site Reliability Engineer improving reliability and deployment speed across enterprise cloud platforms. Automates AWS, Kubernetes (EKS), Terraform, Jenkins, and CI/CD workflows, cutting deployment effort by 85% and reducing operational toil. Builds Prometheus and Grafana monitoring, leads incident response and blameless postmortems, and mentors engineers on operational standards.
AWS
Kubernetes (EKS)
Terraform
Jenkins
Docker
Blue-green deployments
FinOps
Incident management
Root cause analysis
Blameless postmortems
Mentorship
Change management
High availability architecture
Capacity planning
Observability engineering
Disaster recovery planning
Cloud infrastructure design