Summary
Overview
Work History
Education
Skills
Timeline
Generic

ASITAV PATATNAIK

Site Reliability Engineer | Platform Engineer | Cloud Engineer
Bangalore

Summary

Senior Site Reliability, Platform & Cloud Engineer with 8+ years of experience operating and optimizing mission-critical cloud
platforms and production systems at scale, with a proven track record of improving reliability, availability, performance, and
operational efficiency.
Expertise across AWS, Kubernetes, Docker, Terraform, IaC, CI/CD, Linux, Python, Bash, Ansible,GitLab, Datadog, Prometheus,
Grafana, Splunk, and CloudWatch.
Proven ability to reduce MTTR, eliminate operational toil, strengthen production reliability, and automate infrastructure and
operations through observability, incident response, RCA, performance engineering, and disaster recovery.
Strong SRE foundation across SLI/SLO, SLA, error budgets, incident management, resilience, high availability, and production
readiness, with AIOps and AI/ML-assisted operations for intelligent monitoring, anomaly detection, troubleshooting, and
operational automation.

Overview

8
8
years of professional experience

Work History

Senior Site Reliability Engineer

Infosys
10.2021 - Current
  • Led a team in developing cloud-based tooling to automate deployment processes, resulting in a 30% reduction in deployment times for critical services
  • Orchestrated a comprehensive incident response framework, reducing system downtime by 25% year-over-year and improving customer satisfaction ratings.
  • Designed, automated, and supported cloud infrastructure and application deployment platforms across AWS and enterprise production environments, focusing on availability, scalability, security, and operational reliability.
  • Designed and supported containerized workloads using Docker and Kubernetes, including application deployment, scaling, configuration management, health checks, and production troubleshooting.
  • Implemented monitoring, logging, alerting, and operational dashboards using Datadog, CloudWatch, Prometheus, Grafana, and Splunk, improving visibility into infrastructure and application health.
  • Participated in production incident management, troubleshooting, root cause analysis, problem management, and post-incident reviews, driving corrective and preventive actions.
  • Managed a performance testing suite automating CPU and memory benchmarking across production and non-production environments, optimizing resource scaling decisions
  • Provided on-call technical support, ensuring rapid incident resolution and maintaining SLAs for continuous high availability and performance

Senior Software Engineer

IQVIA
06.2020 - 07.2021
  • Designed and implemented CI/CD pipelines to automate build, test, and deployment processes for Clinical Analytics and CBPAS applications hosted on AWS, improving deployment consistency and release reliability.
  • Managed AWS IAM policies, roles, and user groups, applying least-privilege access and security best practices across cloud environments.
  • Developed SQL queries for Oracle and SQL Server to validate production data, investigate application issues, and support incident troubleshooting.
  • Created Grafana dashboards to visualize CPU, memory, disk, network, pod health, application latency, error rates, and service availability.

Senior Software Engineer

CenturyLink
11.2018 - 05.2020
  • Correlated metrics, logs, and alerts across Prometheus, Grafana, Datadog, Splunk, and CloudWatch to improve observability and incident investigation.
  • Developed Ansible playbooks using YAML to automate infrastructure provisioning and development server configuration.
  • Managed containerized environments using Rancher and Docker, including container monitoring and operational troubleshooting.
  • Monitored system health and escalated service risks before they affected availability.
  • Maintained cloud infrastructure components supporting application availability and deployment consistency.

Education

Bachelor of Engineering - Information And Communication Technology

Eastern Academy of Science & Technology
Odisha
05-2011

Skills

  • SLI/SLO/SLA
  • Error Budgets
  • Incident Management
  • Change management
  • Problem Management
  • Jira
  • RCA
  • Postmortems
  • MTTR Reduction
  • High Availability
  • Disaster Recovery
  • AWS
  • EC2
  • ECS
  • EKS
  • Lambda
  • S3
  • RDS
  • ECR
  • VPC
  • IAM
  • CloudWatch
  • Kubernetes
  • Docker
  • EKS
  • Deployments
  • Ansible
  • Terraform
  • GitLab CI/CD
  • Bash
  • Shell Scripting
  • Datadog
  • Prometheus
  • Grafana
  • Splunk
  • CloudWatch
  • AIOps
  • AI/ML-Assisted Operations
  • AM
  • Secrets Management
  • VPC
  • DNS
  • Load Balancing
  • Security Groups
  • TLS/SSL
  • Linux
  • Ubuntu
  • Windows
  • Oracle
  • MySQL
  • Git
  • GitHub
  • GitLab
  • Maven
  • Sonatype Nexus

Timeline

Senior Site Reliability Engineer

Infosys
10.2021 - Current

Senior Software Engineer

IQVIA
06.2020 - 07.2021

Senior Software Engineer

CenturyLink
11.2018 - 05.2020

Bachelor of Engineering - Information And Communication Technology

Eastern Academy of Science & Technology
ASITAV PATATNAIKSite Reliability Engineer | Platform Engineer | Cloud Engineer