t
taukir707

Tauovir Khan

@taukir707

Certified Azure Databricks Data Engineer for Cloud Data Solutions

Canadá
Inglés, Hindi
Parte de la información aparece en idioma inglés.
Sobre mí
Experienced Data Engineer with a strong background in building scalable data solutions across MNCs. Skilled in Azure Databricks, Python, PySpark, and Delta Lake for developing robust ETL pipelines. Experienced in data quality, governance with Unity Catalog, and CI/CD automation using GitHub Actions. Certified as Databricks Data Engineer Associate and Professional. Also proficient in AWS (Glue, Lambda, S3) and Snowflake for cloud-based data engineering.... Lee más

Habilidades

t
taukir707
Tauovir Khan
desconectado • 

Experiencia laboral

Tech_Mahindra

Sr. Data Engineer

Tech Mahindra • Tiempo completo

Jun 2024 - Present2 yrs 2 mos

Working as a Sr. Data Engineer with responsibilities of implementing end-to-end pipeline solutions. Working around Machine learning lifecycle and implemented MLFlow operations on Databricks. Redesigned existing Data pipeline and migrated existing process to modern architecture. Extensively working on Big data technologies Python, PySpark, Spark SQL and Delta lake. 􀀀 Implemented functional and unit tests using Pytest libraries. Maintained Data quality ,Data governance using schema evaluation and Unity Catalog. Orchestrated data pipeline using Databricks workflow. Deployed Databricks Jobs using Databricks Asset bundle. Implemented CI/CD pipeline through Github actions.

Tata_Consultancy Services

Data Engineer

Tata Consultancy Services • Tiempo completo

Oct 2021 - Feb 20242 yrs 4 mos

Worked as a Data Engineer with responsibilities of implementing end-to-end pipeline solutions for data migration projects. Extensively worked on AWS services like AWS glue, Lambda function, S3, SNS, IAM, etc for implementing ETL(extract transform and load) pipeline. Worked on multiple Data Pipeline Frameworks and handled OLAP & OLTP. Worked on Snowflake Data Warehouse and implemented Views, Snowpipe, Tasks, Procedure, etc. Involved in building Dimensional and Data Vault2.0 Models with Data Modeler Team. Extensively used Big data technology PySpark, Python, Pandas, Snowflake and MongoDB. 􀀀 Implemented functional and unit tests using Python Pytest libraries. Developed lake house using Azure Databricks, Azure Data Lake and Delta Lake.