Summary
Overview
Work History
Education
Skills
Personal Information
Certification
LANGUAGES
Timeline
Generic
Debasrita Roy

Debasrita Roy

Kolkata

Summary

Dynamic Technical Architect with extensive experience at Wipro, specializing in cloud architecture and data modeling. Proven track record in optimizing ETL processes and deploying AI models. Adept at collaborating with cross-functional teams to drive innovative solutions, enhancing data governance and operational efficiency. Passionate about leveraging emerging technologies for impactful results.

Overview

1
1
Certification
13
13
years of professional experience

Work History

Technical Architect

Wipro
Bangalore
06.2023 - Current
  • Developed framework to accommodate all source types and establish ETL processes for streamlined data processing and reporting.
  • Configured YAML files using AWS services including S3, Athena, Glue, and Redshift for optimized data storage.
  • Employed AWS Bedrock and SageMaker for comprehensive data analysis.
  • Implemented Agentic AI with a knowledge base to automate repetitive tasks.
  • Collaborated with cross-functional teams to define technical requirements and deliver effective solutions.
  • Utilized Apache Airflow to orchestrate data workflows effectively.

Technical Architect

Wipro
09.2020 - 05.2023
  • Designed scalable architecture for cloud-based applications in diverse client environments.
  • Analyzed requirements and existing environment to devise a strategic system-building approach. Developed Databricks notebooks on Azure, and migrated all previous functionalities seamlessly. Developed Azure Data Factory pipelines to streamline data workflow orchestration. Designed and implemented triggers and scripts to optimize scheduling and enhance performance tuning. Developed a Python framework for handling delta records efficiently. Designed the database schema for optimal data management. Scripted data insertion into Hive tables using Python. Created stored procedures in Synapse for update and merge processes. Collaborated in an Agile development environment, contributing to iterative progress. Estimated project timelines and defined sprint stages for effective project management. Monitored system health and logs, responding to warning or failure conditions to maintain operational stability.
  • Collaborated with cross-functional teams to define technical requirements and solutions.
  • Evaluated emerging technologies to enhance system performance and reliability.

Big Data Lead

TCS
11.2018 - 09.2020
  • Designed and implemented data models for reporting and analytics purposes.
  • Analyzed requirements and existing environment to strategize the SAMs system development. Developed notebooks and workflows in Databricks Azure, ensuring seamless migration of existing functionalities. Configured the Job SLA feature in Databricks Azure to alert on long-running jobs. Created ETL scripts to extract data from RDBMS databases into HDFS, facilitating efficient data processing. Developed Hive tables for data upload from various sources. Designed the database schema for optimal data organization. Scripted data insertion into Hive tables using Python. Stored job statuses in Azure job details for tracking. Proposed and implemented an automated system using Shell script for job sqoop, enhancing data transfer efficiency. Worked within an Agile development framework. Estimated project timelines and defined sprint stages for efficient project management. Developed Hive queries to categorize data for different claims. Monitored system health and logs, responding promptly to warnings or failure conditions.
  • Led data analysis projects for multiple clients across various industries.
  • Collaborated with cross-functional teams to improve data governance processes.

Senior Software Engineer

Accenture
08.2013 - 10.2018
  • Strategy and System Development. Analyzed requirements and existing environment to devise an effective strategy for building the BIC system. Developed Oozie Workflow and Coordinator to integrate Denodo, Hadoop ETL (Hive, Pig, Sqoop), Redshift, and CloudWatch. Enabled Oozie SLA feature to alert on long-running jobs. ETL and Data Integration. Developed ETL scripts to pull data from Denodo database into HDFS. Created Hive tables for uploading data from various sources. Designed and developed scripts to load data from Hive tables into Redshift. Created various views in Redshift for different applications. Stored job statuses in MySQL RDS for tracking. Proposed and implemented automated Shell script for Sqoop jobs, streamlining data transfer processes. Designed distribution styles for Redshift tables to optimize performance. Created shell scripts to copy data from S3 to Redshift, integrating with RDS, Redshift & Hive. ETL Framework Development. Developed and enhanced web-based ETL framework and metadata handler, facilitating efficient data management. Configured, executed, and debugged jobs through the ETL framework. Extracted data from SQL Server, Oracle, Hana, Teradata into HDFS using Sqoop jobs with incremental and full load. Wrote Pig scripts to transform raw data from various sources into baseline data. Developed Hive scripts to perform ad-hoc analysis based on user/analyst requirements. Optimized Hive performance by understanding partitions and designing external tables. Resolved performance issues in Hive and Pig scripts, leveraging knowledge of joins, groups, and aggregations. HBase and Metadata Management. Developed HBase tables for storing metadata to be used by other Hadoop/application-specific elements. Managed full project lifecycle from requirement gathering to deployment.. Owned project lifecycle including requirement gathering, installation, development, unit testing, debugging, and deployment. Key contributor in creating end-to-end data flow through Informatica, including designing and populating the data warehouse and performance tuning. Managed operations for mapping and loading production site data to Oracle data warehouse, ensuring accurate data integration. Handled diverse data sources: SAP, Oracle, MS SQL, Teradata, and flat files. Documentation and Issue Resolution. Authored detailed technical design documents and user test cases. Analyzed issues reported by the onshore team and provided timely solutions by coordinating with offshore developers. Proposed and implemented an automated performance reporting solution using Java.

Education

M.Tech -

Vellore Institute of Technology
01-2013

B.Tech / B.E. -

Narula Institute of Technology
Kolkata
01-2011

XIIth - Science

West Bengal Board of Higher Secondary Education
01-2007

Skills

  • Azure Data Factory
  • AWS Glue ETL
  • Data modeling
  • Data analysis
  • Machine learning
  • AI Model Deployment ML
  • Cloud Architecture
  • Azure Data Lake Storage
  • Data Governance
  • Azure DevOps
  • Azure Functions
  • Serverless Architecture
  • AWS API Management

Personal Information

Title: Technical Architect

Certification

  • CCA Data Analyst
  • AWS Certified Data Engineer - Associate DEA C01

LANGUAGES

Bengali, English, Hindi

Timeline

Technical Architect

Wipro
06.2023 - Current

Technical Architect

Wipro
09.2020 - 05.2023

Big Data Lead

TCS
11.2018 - 09.2020

Senior Software Engineer

Accenture
08.2013 - 10.2018

M.Tech -

Vellore Institute of Technology

B.Tech / B.E. -

Narula Institute of Technology

XIIth - Science

West Bengal Board of Higher Secondary Education
Debasrita Roy