Summary
Overview
Work History
Education
Skills
Designation
declaration
Timeline
Generic

UMASHANKAR V

Bangalore

Summary

Cloud Operations Engineer with 6+ years of experience supporting NOC monitoring & AWS cloud infrastructure, Linux administration production environments, and observability platforms. Experienced in incident management, event monitoring, reliability engineering, root cause analysis, and with a focus on reducing MTTR, Patching activity improving service availability, and strengthening production reliability. Hands-on with Grafana, Prometheus, Alertmanager, AWS services, and modern monitoring frameworks, with a strong interest in Gen AI, MLOps.

Overview

7
7
years of professional experience

Work History

Cloud Operations (NOC Engineer)

Synchronoss Technologies
Bangalore
12.2022 - Current
  • Monitor and manage end-to-end production cloud environments using Prometheus and Grafana to support high availability, performance, and operational visibility.
  • Support AWS services including EKS, EC2, S3, IAM, VPC, EBS, ELB, and CloudWatch across production cloud operations.
  • Administer Linux production environments, including system maintenance, patch management, incident resolution,
  • Provide advanced L2 level production support, diagnosing complex technical issues while meeting SLO and SLA requirements.
  • Use Grafana, Prometheus, Loki, and Alert manager to monitor Kubernetes clusters, cloud infrastructure, APIs, and databases; perform alert triage and troubleshooting.
  • Design and maintain Grafana dashboards, Prometheus alerting rules, exporters, and monitoring strategies to improve visibility and accelerate incident detection.
  • Optimize alerting through threshold tuning and noise reduction to minimize false positives and improve operational effectiveness.
  • Develop and maintain operational runbooks, technical documentation, and Confluence knowledge-base articles to improve operational readiness.
  • Conduct initial root cause analysis (IRCA) and participate in post-incident reviews, implementing preventive measures to reduce recurring incidents.
  • Observability Tool Migration
  • Supported the transition from LogicMonitor to Prometheus and Grafana, modernizing cloud infrastructure monitoring and observability.
  • Evaluated existing dashboards, alert configurations, and monitoring integrations to identify business-critical observability requirements.
  • Developed and maintained custom Grafana dashboards, Prometheus alerting rules, and exporters to improve monitoring coverage and proactive incident detection.
  • Collaborate with Engineering, Infrastructure, DevOps, Platform Engineering, and Application Support teams to resolve critical incidents and improve service availability.
  • Improved infrastructure visibility and monitoring efficiency while reducing monitoring platform costs through adoption of a scalable open-source observability stack.

Network Monitoring

TATA Play Private Limited
Bangalore
06.2019 - 02.2022
  • Monitored network activity 24×7 in Nagios to support system uptime and continuity.
  • Installed and configured requested software packages on Red Hat Enterprise Linux (RHEL 7).
  • Tracked incidents in ServiceNow ticketing tool.
  • Managed P1, P2, and P3 severity issues and collaborated with internal team via bridge calls until services returned to normal.

Education

Intermediate Education -

Vikas Junior College 2010-2012 Batch
Kuppam, Andhra Pradesh

B.Tech - Computer Science Engineering (CSE) 2012-2016 Batch

Jawaharlal Nehru Technological University
Anantapur

Skills

  • ITSM: Incident, Change & Problem Management (ITIL), ServiceNow
  • Cloud & Infrastructure: AWS (EKS, EC2, S3, IAM, VPC, Route53, ELB, CloudWatch)
  • Monitoring & Observability: Prometheus, Grafana, Alertmanager, Loki, LogicMonitor, SolarWinds, AlertSite, Witbe,xMatters Bridge
  • Docker, Argo CD, Git, basics
  • Operating Systems: Linux (RHEL 7, Ubuntu, CentOS), Windows
  • Collaboration: Confluence, Microsoft Teams, Zoom

Designation

CLOUD OPERATIONS ENGINEER (NOC)

declaration

I hereby declare that all the information provided by me in the above context is true to the best of my knowledge.

Timeline

Cloud Operations (NOC Engineer)

Synchronoss Technologies
12.2022 - Current

Network Monitoring

TATA Play Private Limited
06.2019 - 02.2022

Intermediate Education -

Vikas Junior College 2010-2012 Batch

B.Tech - Computer Science Engineering (CSE) 2012-2016 Batch

Jawaharlal Nehru Technological University
UMASHANKAR V