

Senior System Operations Engineer improving production stability for 5–8 critical equipment finance and KYC applications. Delivers higher uptime, fewer recurring incidents, and faster detection through automation that saves 10–20 hours per month, structured problem management, and observability. Partners with platform and application teams on 5–8 off-hours releases per month, OpenShift migration, and secure cloud operations across Azure and Linux-based environments.
AWS
Microsoft Azure
Docker
Jenkins
Prometheus
Platform reliability engineering
Incident management
Change management
Release management
Disaster recovery planning
Team leadership
Operational governance
Mentorship
Capacity management
Stakeholder communication