Dynamic DevOps and Cloud Infrastructure professional with a proven track record in designing, implementing, and maintaining robust systems. Effective management of teams and streamlined documentation processes ensure operational excellence. Expertise includes IaaS, PaaS, and SaaS solutions, with a strong focus on optimizing DevOps, CloudOps, and Infrastructure Management practices. Committed to driving innovation and efficiency while seeking new challenges within a growth-oriented organization that values collaboration and forward-thinking strategies.
Overview
13
13
years of professional experience
4021
4021
years of post-secondary education
Work History
Senior Manager DevOps and Cloud-Infra Engineer
Crypto.com
Bengaluru, India (Remote)
11.2023 - Current
Led a comprehensive cloud migration initiative, successfully transitioning all on-premise infrastructure and services to the AWS cloud. This involved orchestrating the migration of critical environments (including DCO, DCM, FCM, and the OEX Exchange) by leveraging a cloud-native migration strategy, Infrastructure as Code (IaC), and CI/CD pipelines, to ensure a seamless and efficient transition.
Architected and deployed microservices architectures, utilizing Nginx and Kong as API gateways and reverse proxies to manage traffic, enhance security, and improve system modularity.
Spearheaded a GitOps transformation by implementing ArgoCD and Kustomize to manage application deployments, ensuring declarative configuration management, and automated synchronization with Git repositories.
Designed and implemented end-to-end CI/CD pipelines using GitHub Actions, Jenkins, Azure DevOps, Bamboo and Bitbucket, significantly reducing deployment times and streamlining development workflows.
Implemented containerization using Docker and Kubernetes, facilitating seamless deployment and application scalability.
Created and managed Kubernetes clusters, including monitoring and logging setups, ensuring high availability, and performance.
Developed and maintained Infrastructure as Code (IaC) using Terraform, provisioning and managing cloud resources, including VMs, networks, and storage, for efficient and repeatable deployments.
Managed and optimized cloud infrastructure on Azure and AWS for on-premise solutions.
Automated configuration management with Ansible enhances system consistency and reliability.
Leveraged Strimzi to implement Kafka MirrorMaker for seamless data replication between source and target clusters, ensuring business continuity, and disaster recovery capabilities.
Automated the configuration and deployment of Kafka components, including MirrorMaker, using Infrastructure as Code (IaC) principles with YAML and Kubernetes Custom Resources, reducing manual effort and potential errors.
Implemented and managed an in-house Kafka cluster, including MSK and Kafbat, to handle high-throughput data streams.
Configured data serialization for Kafka using Avro and binary formats, optimizing data transfer, and ensuring data integrity.
Established a comprehensive monitoring and alerting stack using Prometheus and Grafana, providing real-time visibility into system health and performance.
Implemented comprehensive monitoring and alerting systems using Grafana and Loki, improving system observability.
Implemented a two-way synchronization feature between Grafana and Git, enabling version control and collaborative management of dashboards.
Engineered robust security measures by configuring Palo Alto firewalls and Azure Network Security Groups (NSGs), protecting against cyber threats, and ensuring compliance.
Configured Kong as an API Gateway to provide centralized authentication, security, and traffic management for microservices.
Leveraged Kustomize to manage environment-specific configurations for Kubernetes applications, maintaining a centralized base while using overlays for streamlined deployments across development, staging, and production environments.
Utilized Git and Bitbucket for version control, maintaining code integrity, and facilitating collaboration.
Developed and deployed a microservices architecture, improving system modularity and maintainability.
Azure PaaS, like App Services, Web Apps, App Service Plans, and environments.
Load balancing using HAProxy, Load Balancer, and Traffic Manager.
AKS, Infra setup, and CICD pipelines for APIs/UIs.
Team leadership, business strategy, team motivation, Sitecore.
Helm Charts and Ingress Controller Setup.
Handled configuration management using the best-suited software tools, including Terraform, CloudFormation, and Ansible.
Senior Associate
Goldman Sachs
Bengaluru
02.2015 - 03.2022
Implement the latest technologies, various monitoring and remediation tools, vendor connectivity with various market data feeds, and complex network systems to improve effectiveness.
Working on AWS services: VPC, ECS, EC2, RDS, S3, CloudWatch, IAM, SNS, Route 53, Autoscaling, ELB, WAF, Lambda, etc.
Moving applications from legacy VPN to AWS Direct Connect.
Setting up the KONG reverse proxy for applications from GSINET to the B2B DMZ.
Moving applications from legacy Dashrouters to the KONG structure.
Internal and external site certificates should be updated in a timely manner, and a check should be kept on their expiry.
Automate business health checks, numerous manual jobs, user login, and user configurations to better utilize cost and effort.
Setting up SAML and SSO authentication for various vendor services.
Setting up internal Autosys jobs for production servers for timely feeds.
Switched the business to GIT version control, building, and supporting CI/CD.
Assisted with refining the Software Development Life Cycle to fit system needs.
I wrote Terraform scripts to automatically update system components, saving 30% of admin time.
Appproxy setup via Conman and Conduit.
Managing multiple Linux servers for various internal apps and services.
Working with the TechRisk team to ensure compliance with centrally defined security.
Monitor scheduled power shutdown activities, pre-packaging testing process, code releases, quality checks, firewall/infrastructure security, auction system setup, outage issues, etc.
Planning and building strategies for new and upcoming projects and changes to drive them successfully, accordingly raising GCM (Global Change Management Tickets) to cover each part of the activity.
Testing, onboarding of trading applications in the firm via proper channels, working with the packaging team, and then rolling it out to the audience in phases.
Conducting training for internal engineers, and creating product-based documentation on Confluence.
Maintaining, onboarding, and administering web-based third-party apps on the internal portal called GSACS, such as Tradeweb, MarketAxess, FastTrack, and XTAuction Italy/Spain, to name a few.
Updating the database of the applications and users on the App-share.
Testing and updating of applications launch scripts on respective servers at the time of change, and maintaining the database of the same.
Working with the vendor on escalated firm-wide issues and onboarding applications.
Raising Sev's, driving it to closures with RCA, and sending notifications to the impacted audience.
Senior Analyst
HCL Technologies
Noida
05.2013 - 01.2015
The installation, configuration of components, and troubleshooting of Citrix XenApp.
Application installation, publication, and isolation within the Citrix environment.
Support client applications installed on Citrix.
Troubleshooting and fixing preliminary issues related to the Citrix Web Interface, Citrix Receiver, etc.
Installing/configuring Windows updates as per existing policies.
Troubleshooting on Windows Server 2008, Delegate Control, and Shadow Copies.
Provisioning and de-provisioning users' access in AD and other applications.
Handling users' issues related to AD.
Preparing an incident report for SEV1 and SEV2 issues.
Analyst
Wipro
New Delhi
09.2012 - 03.2013
All technical troubleshooting of the Windows operating system is done by taking a remote session of the client's system.
Technical query handling.
Troubleshooting/Connecting to the Network Printer.
Severity Handling: Provide on-call support for Severity 1 issues, and escalation handling.
Key Roles in Severity Handling.
Accountability in Severity Handling (Incident and Problem Management):
Problem Management: Manage problems to ensure that these are diagnosed, logged, and escalated to appropriate and consistent quality standards.
Trend Analysis: Produce a trends analysis of recurring problems/incidents, extracting trends on incident types, customer types, key problem areas, departments, hardware types, etc.
Customer Interface: Delivering and managing high-standard communications across customers and IT to ensure that problems are dealt with by priority and customer needs.