
Over all 5 Years of IT experience on data driven application design and development with 3.5 years of relevant hands-on-experience in various Azure data services and SQL databases. Dedicated and highly skilled Azure Data Engineer with an experience in designing, implementing and managing data solutions on the Microsoft Azure platform. Proven expertise in utilizing Azure Databricks, Azure Data Factory (ADF) and Azure Data Lake Storage (ADLS), SQL, Pyspark.
Project1: Prudential
Client: Prudential Singapore
Environment: Azure Data Bricks, Pyspark, SQL Database, ADLS Gen2, ADF, Datawarehouse, Data Lake, GraphQL, logic Apps
Description: Prudential Assurance Company Singapore (PACS) is one of Singapore’s leading life insurance companies. PACS is transitioning to become a more innovative and agile organization that can adapt quickly to the rapidly changing needs of our customers.
Roles & Responsibilities
• Designing and implementing highly performant data pipelines from multiple sources using Apache Spark and/or Azure Databricks • Integrating the end-to-end data pipeline to take data from source systems to target data repositories ensuring the quality and consistency of data is always maintained
• Build a Spark Streaming pipeline with Synapse.
• Developed Pyspark code to read, transform, and write data for Batch Processing
• Designing Spark Cluster for a Streaming ETL Process
• Build ETL Script to implement Spark Structured Streaming that guarantees exactly-once stream processing
• Designing Spark Clusters for a Streaming ETL Process
• Created Data frames in Databricks and applied various transformations like string functions, aggregations, window functions, Filtering, Splitting, Renaming, Removing duplicates etc.
Project2: Manulife – VST
Client: Ameriprise Financial
Environment: Azure data factory, SQL Server, PowerBI, Pyspark, Azure data bricks
Roles & Responsibilities
• Worked on building the data pipeline using Azure Service like Data Factory to load the data from Legacy, SQL server to Azure Data warehouse using Data Factories and Databricks Notebooks.
• Build complex ETL jobs that transform data visually with data flows or by using compute services Azure Databricks and Azure SQL Database.
• PowerBI tool is used for create interactive visualizations and reports from data. Its connected directly to Azure SQL database or Azure Analysis Services to fetch data for Visualization.
• Use various types of activities: data movement activities, transformations and control activities; Copy data, Data flow, Get Metadata, Lookup, Stored procedure, Execute Pipeline.
• Created Pipeline’s to extract data from on premises source systems to azure cloud data lake storage.
• Executed Extract, Transform, and Load (ETL) processes, transferring data from source systems to Azure
Data Storage service using Azure Data Factory.
DOB : 30/09/1993
Address : Pune, MH, India
Available : Immediate Joiner
www.linkedin.com/in/akash-tijare