Summary
Overview
Work History
Education
Skills
Accomplishments
Technical Environment
Core Technical Skills
Timeline
Generic

PRATIK KUMAR AGASTI

Mumbai

Summary

Results-driven Data Center Operations Lead with over 12 years of expertise in large-scale data center management and infrastructure deployment. Skilled in supporting air-cooled and liquid-cooled environments, incident management, and operational readiness. Demonstrated track record in managing significant server deployments and resolving infrastructure incidents, while leading technical teams to execute large-scale projects effectively.

Overview

12
12
years of professional experience

Work History

Data Center Operations (DCO) Lead

AMAZON WEB SERVICES (AWS)
2022.06 - Current
  • Led large-scale data center build, infrastructure deployment, operations, project execution, technical team leadership, vendor coordination, and operational readiness.
  • Lead large-scale data center build and infrastructure deployment activities across both air-cooled and liquid-cooled environments.
  • Lead and coordinate a team of 50+ technical resources/builders, including work allocation, shift planning, task assignment, and operational execution.
  • Manage multiple data center projects from planning and execution through completion and operational handover.
  • Directed installation, deployment, and troubleshooting activities for 5,000+ data center servers, supporting infrastructure reliability and availability.
  • Supported data center construction, infrastructure deployment, commissioning, operational readiness, and daily operations to ensure seamless functionality.
  • Managed vendors and contractors to meet project milestones and uphold quality standards, safety requirements, and operational objectives.
  • Handled critical data center escalations and coordinated cross-functional teams for timely resolution of infrastructure issues.
  • Perform root cause analysis (RCA) and technical deep dives for server, storage, network, rack, and infrastructure-related issues.
  • Configure, install, troubleshoot, and maintain Dell, HPE/HP, and Lenovo servers and associated infrastructure.
  • Perform hands-on troubleshooting of server hardware, storage infrastructure, network devices, racks, cabling, and connectivity.
  • Support large-scale environments consisting of diverse compute, host, storage, and network infrastructure.
  • Build and support high-density and Machine Learning (ML) data center environments.
  • Deployed data center infrastructure components for liquid-cooled environments to ensure successful setup and operational activities.
  • Monitor data center infrastructure using various monitoring, deployment, and operational management tools.
  • Respond to critical infrastructure events involving servers, storage, network devices, racks, and associated systems.
  • Provide Tier 2 escalation support for complex server hardware and network-related incidents, including 1,000+ locally managed incidents.
  • Work closely with Server, Storage, Network, Facilities, Safety, Security, and Project teams to maintain operational continuity.
  • Support CapEx and OpEx planning, resource allocation, project execution, and operational budgeting.
  • Supported technical training, knowledge development, and recruitment as a certified instructor and interviewer.
  • Created shift rosters to manage workforce effectively and address service demands., allocate daily work, prioritize critical activities, and ensure adequate technical coverage.
  • Maintain strong focus on safety, security, operational excellence, infrastructure reliability, and continuous improvement.

System Support Engineer – Payroll: TPM Guru Pvt. Ltd.

RELIANCE COMMUNICATIONS
2021.04 - 2022.06
  • Managed and supported enterprise storage platforms including EMC VNX 5300/5400/5500, CLARiiON CX-300, Unity 350, IBM DS3200/4300/4700, HP EVA 4400/8100, and HP MSA 1000/2000.
  • Performed storage provisioning activities, including creation and management of RAID groups, storage pools, logical drives, and storage allocations.
  • Allocated storage resources aligned with Windows and Linux server requirements and optimized for storage tiers.
  • Administered SAN fabric environments using Brocade switches, including hard zoning and soft zoning.
  • Troubleshot SAN connectivity and storage issues, ensuring reliable host-to-storage connections and addressing multipathing challenges.
  • Supported Linux/Unix server environments, including installation, configuration, troubleshooting, LVM, SVM, and multipathing.
  • Performed hardware fault diagnosis and log analysis using HP SmartStart, Live OMSA, and DSET.
  • Troubleshot hardware issues on HP ProLiant and Dell PowerEdge servers.
  • Executed server hardware replacement, conducted fault isolation, performed health checks, and validated systems post-maintenance.

Customer Support Engineer – Payroll: TERiX International

TATA TELE BUSINESS SERVICES
2020.05 - 2021.03
  • Provided infrastructure and data center hardware support for HP ProLiant BL/DL/ML series and Dell PowerEdge Blade Servers.
  • Installed, configured, maintained, and troubleshot enterprise server hardware to ensure system reliability and performance.
  • Supported Sun Netra and Sun Fire server platforms, including models 210, 440, 4800, and M/X series.
  • Provided hardware support for enterprise storage platforms including Sun StorageTek 2540/3500/6140, EMC, HP EVA, and Dell PowerVault MD3000.
  • Diagnosed hardware faults, replaced components, performed preventive maintenance, and conducted infrastructure health checks to minimize downtime.
  • Supported data center hardware incidents and coordinated with internal teams and vendors to resolve critical issues.
  • Maintained operational documentation and supported daily infrastructure activities to facilitate seamless operations.

Senior Customer Support Engineer

IND INNOVATION PVT. LTD.
2014.12 - 2020.04
  • Installed, configured, maintained, and troubleshot Linux servers, including RHEL 5, RHEL 6, and RHEL 7.
  • Troubleshot server performance and hardware issues related to RAID, CPU, memory, file systems, and other critical components.
  • Monitored infrastructure health and escalated complex incidents to technical teams, facilitating swift resolution and minimizing downtime.
  • Provided onsite and remote data center infrastructure support, including server hardware, operating systems, and system administration.
  • Managed day-to-day data center operations, maintaining infrastructure availability and stability for uninterrupted service.
  • Performed hardware troubleshooting, break/fix activities, component replacement, and preventive maintenance.
  • Conducted daily health checks of SAN switches, network switches, and routers to ensure optimal performance and reliability.
  • Performed rack, tower, and blade server installation, mounting, cabling, and commissioning.
  • Generated and reviewed storage pool capacity and utilization reports.

Education

Bachelor of Technology (B.Tech) - Computer Science & Engineering

Biju Patnaik University of Technology
Odisha

Diploma - Computer Science & Engineering

Dhabaleswar Institute of Polytechnic
Odisha

10th Board Examination -

Board of Secondary Education
Odisha

Skills

  • Data center operations
  • Site rollout
  • Server deployment
  • Storage Administration & Provisioning
  • SAN & SAN Fabric Management
  • Brocade SAN Switches
  • RAID management
  • Linux Administration
  • Storage Management
  • Storage Virtualization
  • Data Redundancy
  • Network Infrastructure Deployment
  • Network Switches & Routers
  • Rack and Blade Server Installation
  • Server Hardware Troubleshooting
  • Hardware repairs
  • Data Center Monitoring
  • Incident management
  • Root cause analysis
  • Event management
  • Project management
  • Vendor management
  • Team leadership

Accomplishments

  • Currently working as a Data Center Operations (DCO) Lead at Amazon Web Services (AWS), with hands-on experience supporting large-scale air-cooled and liquid-cooled data center environments.
  • Supported and directed deployment and troubleshooting activities for 10,000+ data center servers.
  • Resolved and managed 1,000+ server hardware and network infrastructure incidents as a Tier 2 escalation point.
  • Lead and coordinated 50+ technical resources/builders in large-scale data center operations and build activities.
  • Supported the rollout of 20 data halls and an ML Operations Center with approximately 2000+ racks as part of the AWS AP-South-1C site project.
  • Hands-on experience across compute, storage, SAN, network, and data center infrastructure.
  • Experience supporting both air-cooled and liquid-cooled data center environments.
  • Extensive experience with enterprise infrastructure from Dell, HPE/HP, Lenovo, EMC, IBM, Sun/Oracle, Dell PowerVault, and Brocade.
  • Experienced in critical-event management, infrastructure troubleshooting, RCA, vendor coordination, and operational readiness.

Technical Environment

Dell PowerEdge, HPE/HP ProLiant, Lenovo, Sun/Oracle Netra, Sun Fire, Rack Servers, Tower Servers, Blade Servers, EMC VNX 5300/5400/5500, EMC CLARiiON CX-300, EMC Unity 350, IBM DS3200/4300/4700, HP EVA 4400/8100, HP MSA 1000/2000, Sun StorageTek 2540/3500/6140, Dell PowerVault MD3000, Brocade SAN Switches, SAN Fabric, Hard Zoning, Soft Zoning, Storage Provisioning, RAID, Storage Pools, Multipathing, Linux/Unix, Red Hat Enterprise Linux (RHEL) 5/6/7, Windows Server, Network Switches, Routers, Network Interfaces, Console Infrastructure, iLO, Network Connectivity Troubleshooting, Data Center Build, Server Deployment, Hardware Break/Fix, Rack & Cabling, Infrastructure Monitoring, Incident Management, Critical Event Management, Root Cause Analysis, Preventive Maintenance, Team Leadership, Vendor Management, Project Management, Workforce Planning, Shift Management, CapEx/OpEx Planning, Safety & Security, Technical Training, Interviewing

Core Technical Skills

  • Data Center Operations & Infrastructure
  • Data Center Build & Site Rollout
  • Air-Cooled & Liquid-Cooled Data Centers
  • Server Installation & Deployment
  • Server Hardware Troubleshooting
  • ML Server Hardware Troubleshooting
  • Storage Administration & Provisioning
  • SAN & SAN Fabric Management
  • Brocade SAN Switches
  • RAID & Storage Pools
  • Linux Administration
  • LVM, SVM & Multipathing
  • Network Switches & Routers
  • Network Infrastructure Deployment
  • Rack & Blade Server Installation
  • Hardware Break/Fix & Component Replacement
  • Data Center Monitoring
  • Incident & Escalation Management
  • Root Cause Analysis (RCA)
  • Critical Event Management
  • Vendor & Contractor Management
  • Project Management
  • CapEx & OpEx Planning
  • Team Leadership
  • Workforce & Shift Planning
  • Safety & Security Compliance
  • Technical Training & Interviewing

Timeline

Data Center Operations (DCO) Lead

AMAZON WEB SERVICES (AWS)
2022.06 - Current

System Support Engineer – Payroll: TPM Guru Pvt. Ltd.

RELIANCE COMMUNICATIONS
2021.04 - 2022.06

Customer Support Engineer – Payroll: TERiX International

TATA TELE BUSINESS SERVICES
2020.05 - 2021.03

Senior Customer Support Engineer

IND INNOVATION PVT. LTD.
2014.12 - 2020.04

Bachelor of Technology (B.Tech) - Computer Science & Engineering

Biju Patnaik University of Technology

Diploma - Computer Science & Engineering

Dhabaleswar Institute of Polytechnic

10th Board Examination -

Board of Secondary Education
PRATIK KUMAR AGASTI