

Skilled Azure Data Engineer with 6 years of experience in designing and implementing scalable data solutions using Azure Databricks, PySpark, SQL, and Azure Data Factory. Specializes in building high-performance ETL pipelines for structured and semi-structured datasets and implementing Delta Lake architectures in cloud environments. Focused on distributed data processing and transformation to deliver reliable analytics and data platforms.
Developed scalable ETL pipelines with Databricks and PySpark to streamline data ingestion and processing.
Implemented Delta Lake architecture to enhance data processing reliability and optimization.
Built Azure Data Factory pipelines to automate orchestration and scheduling of data workflows.
Processed large datasets from Azure Data Lake Storage (ADLS Gen2).
Wrote optimized SQL queries to enhance data transformation and support analytics needs.
Improved Spark job performance through partitioning, caching, and broadcast joins.
Azure Databricks and PySpark
ETL pipelines and data transformation
SQL and Azure Data Factory
Data Lake Storage (ADLS Gen2) and Delta Lake
Distributed data processing
Power BI analytics
Data governance
Agile development methodologies