Summary
Overview
Work History
Education
Skills
Certification
Certifications & Training
Timeline
SATHIESH

SATHIESH

Systems Engineer NOC and Production Support
Bengaluru,KA

Summary

NOC and Production Support professional with 6+ years of experience providing NOC, Production and Non-Production Support across IT infrastructure and application environments. Experienced in Major Incident Management, Problem Management and Change Management, with a proactive and reactive approach to identifying operational risks, preventing recurring incidents, resolving production issues and coordinating controlled changes. Strong background in high-severity incident response, monitoring, troubleshooting, service-impact assessment, SLA management, stakeholder communication and cross-functional coordination. Hands-on experience with Azure Portal, AKS monitoring, Splunk, Splunk Observability, Check_MK, AppDynamics, PRTG, SolarWinds, Park My Cloud, IBM Turbonomic and ServiceNow. Developed operational automations using Python, PowerShell and Power Automate, including Splunk log reporting, multi-server service checks, Excel/OneDrive data processing and recurring operational workflows. Experienced in developing React/HTML-based operational health dashboards to improve application and infrastructure health visibility and reporting. Strong in structured troubleshooting, proactive monitoring, reactive incident response, root-cause-oriented problem management, change coordination, technical documentation and continuous improvement within 24x7 NOC and production support environments.

Overview

2
2
Certifications
6
6
years of professional experience

Work History

Systems Engineer – NOC / Incident Management

INFOSYS
04.2022 - Current
  • Manage Major Incidents and P1/P2 incidents, assessing business impact and urgency, coordinating resolver teams, leading incident bridge calls and driving timely restoration of services while maintaining SLA commitments.
  • Apply both proactive and reactive approaches to Incident and Problem Management, responding to service-impacting issues while identifying recurring patterns and opportunities to prevent future incidents.
  • Perform Problem Management activities by analyzing recurring incidents, reviewing operational symptoms and monitoring data, supporting root-cause investigations and coordinating corrective or preventive actions with resolver teams.
  • Performed problem management by analyzing recurring incidents, reviewing operational symptoms and monitoring data, supporting root-cause investigations, and coordinating corrective or preventive actions with resolver teams to enhance service reliability.
  • Monitor infrastructure and application environments using Splunk, Splunk Observability, Check_MK, AppDynamics, PRTG, SolarWinds, Park My Cloud and IBM Turbonomic, investigating alerts, performance indicators and service-health conditions.
  • Perform Azure Portal monitoring and preliminary troubleshooting across production and non-production environments, including alert review, log analysis and resource-level investigation; escalate issues requiring elevated access to the Azure support team.
  • Manage incidents, problems and changes through ServiceNow, maintaining accurate operational records, escalation details, implementation information and stakeholder communications.
  • Analyzed monitoring information, application/infrastructure alerts, and exported Splunk logs to support incident triage, troubleshoot issues, assess service impact, and conduct problem analysis for operational reporting.
  • Provide NOC, Production and Non-Production Support for infrastructure and application environments within a 24x7 operational support model.
  • Perform Change Management activities across production and non-production environments, coordinating deployments, patching, scheduled maintenance and other controlled changes in accordance with approval, impact and implementation requirements.
  • Apply a proactive approach to Change Management by assessing potential service impact, dependencies and operational risks before implementation, while taking a reactive approach when emergency changes are required to restore or protect services.
  • Prepare incident reports, service-impact communications, problem-related analysis and productivity reports while supporting continuous improvement of operational processes.
  • Developed PowerShell automation to check and start selected or multiple Windows services across different servers from a jump server, supporting repetitive operational service-management activities.
  • Developed React-based health check dashboard to modernize daily collection of application and infrastructure health status, replacing Excel and improving accessibility of critical information.
  • Implemented searchable health-status views, RED/AMBER/GREEN categorization and CSV export functionality for operational and management reporting.
  • Developed an HTML-based PULSE dashboard prototype covering application health, critical jobs, major issues, deployment status and nightly refreshes using email-derived information and exported JSON data.
  • Identify repetitive operational activities, recurring incidents and manual processes suitable for automation or process improvement to improve operational efficiency and service reliability.

Project Engineer

SYMBIOTIC AUTOMATION SYSTEMS PVT. LTD.
06.2020 - 03.2022
  • Troubleshot industrial automation systems to resolve technical issues, enhancing client system reliability.
  • Investigated system issues and implemented corrective actions while supporting project delivery and technical quality.
  • Collaborated with cross-functional teams to resolve automation-related issues and support successful client project execution.
  • Collaborated with cross-functional teams to address automation-related issues, facilitating timely project delivery.

Education

B.Tech. - Artificial Intelligence and Machine Learning

Birla Institute of Technology & Science, Pilani, Pilani
12-2028

Diploma - Electronics and Communication

MN Technical Institute, Bengaluru, India
12-2019

Skills

NOC Support

Production Support

Non-Production Support

Major Incident Management

Incident Management

Problem Management

Change Management

Proactive Problem Management

Reactive Incident Response

Proactive Change Management

Reactive Change Management

Incident Triage

Troubleshooting

Root Cause Analysis

Escalation

SLA Management

Deployment Support

Patch & Maintenance Support

Health Checks

Splunk

Splunk Observability

Check_MK

AppDynamics

PRTG

SolarWinds

Park My Cloud

IBM Turbonomic

ServiceNow

Log Analysis

Alert Investigation

Infrastructure Monitoring

Application Monitoring

Capacity/Performance Monitoring

Operational Reporting

Python

PowerShell

Power Automate

React

HTML

CSV Reporting

Excel/OneDrive Automation

Operational Workflow Automation

Microsoft Azure

AKS Monitoring

Azure Alerts

Azure Logs

Resource Troubleshooting

Windows Server

Networking Fundamentals

Ubuntu/Linux

SSH

UFW

Docker

Portainer

Git/GitHub

MySQL Fundamentals

SQL Fundamentals

Cybersecurity/SOC Fundamentals

Certification

Microsoft Certified: Azure Fundamentals (AZ-900)

Certifications & Training

  • ITIL Awareness
  • Data Privacy Compliance
  • Agile Operations
  • Privacy by Design
  • Cybersecurity in SDLC
  • Python
  • SQL
  • PowerShell
  • Networking
  • Windows Administration
  • ServiceNow

Timeline

Systems Engineer – NOC / Incident Management - INFOSYS
04.2022 - Current
Project Engineer - SYMBIOTIC AUTOMATION SYSTEMS PVT. LTD.
06.2020 - 03.2022
Birla Institute of Technology & Science, Pilani - B.Tech., Artificial Intelligence and Machine Learning
MN Technical Institute - Diploma, Electronics and Communication
SATHIESH Systems Engineer NOC and Production Support