Summary
Overview
Work History
Education
Skills
Websites
Certification
Timeline
Core Competencies
Generic

JITENDRA KUMAR

Bengaluru

Summary

Senior Data Engineer with 9 years of experience in ETL development, data integration, warehousing, and cloud (AWS & Azure). Expertise in designing and optimizing large-scale data pipelines, improving processing efficiency (up to 40%), and ensuring data quality and governance compliance. Skilled in Apache Spark, PySpark, Talend, Informatica, SQL, Python, Scala, and Java. Experienced in AWS (S3, Glue, Redshift, Lambda, EMR, Athena) and Azure (ADF, Blob Storage, Databricks, Synapse). Proven ability to lead teams, mentor engineers, and deliver data-driven business solutions across Banking, Finance, Telecom, and Marketing domains.

Overview

9
9
years of professional experience
1
1
Certification

Work History

Senior Data Engineer

ANZ
Bengaluru
12.2023 - Current
  • Designed end-to-end ETL pipelines (Apache Spark, Python, Hive) reducing data processing time by 35%.
  • Integrated datasets from CRM, online ad platforms, and sales systems, improving data accuracy by 22%.
  • Automated workflows using Airflow, achieving 99.9% SLA adherence.
  • Implemented data governance practices (lineage, validation checks) for compliance.
  • Projects: Portfolio Analytics, Commercial Campaign | Team Size: 5
  • Technologies: Spark, Python, SQL, Hive, Teradata, Starburst, Airflow, AWS S3, GitHub, Jupyter

Senior Data Engineer

Altimetrik
Bengaluru
04.2021 - 11.2023
  • Migrated DataMarts & Data Lakes to AWS Cloud (S3, Glue, Redshift, EMR, Lambda, Step Functions) reducing infra cost by 20%.
  • Built real-time pipelines (Spark, Python, Scala) reducing latency from 8 hours to 45 minutes.
  • Developed Talend workflows improving data quality scores by 30%.
  • Mentored and managed a team of 10 engineers, improving delivery efficiency by 15%.
  • Projects: Brand Marketing, Developer & Partner, DATA GetWell, Finance Data
  • Technologies: Spark, Python, Scala, Talend, Informatica, AWS Glue, Redshift, Lambda, Databricks, Tidal

Senior Software Engineer

Altran – Amdocs
Bengaluru
06.2019 - 03.2021
  • Built Azure Data Factory (ADF) & Databricks pipelines cutting processing time by 40%.
  • Designed models in Azure Synapse & Blob Storage increasing query performance by 28%.
  • Automated monitoring via Azure Functions, reducing error resolution time by 50%.
  • Projects: Vodafone India, Century Link, VEON, Kyiv Star, SingTel | Team Size: 6–8
  • Technologies: PySpark, Hive, SQL, ADF, Databricks, Snowflake, Azure Blob Storage

Implementation Engineer

6D Technologies
Bengaluru
10.2016 - 06.2019
  • Built Hadoop & Spark pipelines for telecom billing/customer data.
  • Automated ETL workflows using Python, Sqoop, Shell scripting, reducing manual work by 60%.
  • Scheduled jobs with Oozie & crontab, ensuring high availability.
  • Delivered data quality frameworks and QA standards for compliance.
  • Projects: APUA, Globecom, NEDDA, Ooredoo | Team Size: 8
  • Technologies: Spark, Hive, Sqoop, Shell, Linux, Hadoop

Education

B.E. - Computer Science

Visvesvaraya Technological University
01.2016

Skills

  • ETL Tools: Apache Spark, Talend, Informatica, Hive, Sqoop
  • Cloud: AWS (S3, Glue, EMR, Redshift, Lambda, Athena), Azure (ADF, Databricks, Synapse, Blob Storage)
  • Programming: Python, SQL, Scala, Java, Shell Scripting
  • Data Governance: Data Lineage, Metadata Management, Quality Frameworks
  • Orchestration: Airflow, Oozie, Skybot, Tidal
  • Version Control & Tools: GitHub, JIRA, Jenkins, Databricks, Snowflake

Certification

  • AWS Certified Data Analytics – Specialty
  • Microsoft Certified: Azure Data Engineer Associate
  • Databricks Certified Data Engineer Associate

Timeline

Senior Data Engineer

ANZ
12.2023 - Current

Senior Data Engineer

Altimetrik
04.2021 - 11.2023

Senior Software Engineer

Altran – Amdocs
06.2019 - 03.2021

Implementation Engineer

6D Technologies
10.2016 - 06.2019

B.E. - Computer Science

Visvesvaraya Technological University

Core Competencies

  • ETL Development & Data Pipelines: Apache Spark, PySpark, Talend, Informatica, Hive, Sqoop
  • Cloud Platforms: AWS (S3, Glue, EMR, Redshift, Lambda, Athena), Azure (ADF, Databricks, Synapse, Blob Storage)
  • Programming: Python, SQL, Scala, Java, Shell Scripting
  • Big Data & Warehousing: Hadoop, Teradata, DBT, Starburst (Presto), Snowflake
  • Data Governance & Quality: Lineage, Metadata Management, Validation Frameworks
  • Orchestration Tools: Airflow, Oozie, Tidal, Skybot
  • Project Leadership: Agile/Scrum, Estimation, Stakeholder Collaboration
  • Tools: GitHub, Jenkins, JIRA, Databricks, Alation
JITENDRA KUMAR