t
taukir707

Tauovir Khan

@taukir707

Certified Azure Databricks Data Engineer for Cloud Data Solutions

Canada
Engels, Hindi
Sommige informatie wordt in het Engels weergegeven.
Over mij
Experienced Data Engineer with a strong background in building scalable data solutions across MNCs. Skilled in Azure Databricks, Python, PySpark, and Delta Lake for developing robust ETL pipelines. Experienced in data quality, governance with Unity Catalog, and CI/CD automation using GitHub Actions. Certified as Databricks Data Engineer Associate and Professional. Also proficient in AWS (Glue, Lambda, S3) and Snowflake for cloud-based data engineering.... Lees meer

Skills

t
taukir707
Tauovir Khan
offline • 

Werkervaring

Tech_Mahindra

Sr. Data Engineer

Tech Mahindra • Fulltime

Jun 2024 - Present2 yrs 2 mos

Working as a Sr. Data Engineer with responsibilities of implementing end-to-end pipeline solutions. Working around Machine learning lifecycle and implemented MLFlow operations on Databricks. Redesigned existing Data pipeline and migrated existing process to modern architecture. Extensively working on Big data technologies Python, PySpark, Spark SQL and Delta lake. 􀀀 Implemented functional and unit tests using Pytest libraries. Maintained Data quality ,Data governance using schema evaluation and Unity Catalog. Orchestrated data pipeline using Databricks workflow. Deployed Databricks Jobs using Databricks Asset bundle. Implemented CI/CD pipeline through Github actions.

Tata_Consultancy Services

Data Engineer

Tata Consultancy Services • Fulltime

Oct 2021 - Feb 20242 yrs 4 mos

Worked as a Data Engineer with responsibilities of implementing end-to-end pipeline solutions for data migration projects. Extensively worked on AWS services like AWS glue, Lambda function, S3, SNS, IAM, etc for implementing ETL(extract transform and load) pipeline. Worked on multiple Data Pipeline Frameworks and handled OLAP & OLTP. Worked on Snowflake Data Warehouse and implemented Views, Snowpipe, Tasks, Procedure, etc. Involved in building Dimensional and Data Vault2.0 Models with Data Modeler Team. Extensively used Big data technology PySpark, Python, Pandas, Snowflake and MongoDB. 􀀀 Implemented functional and unit tests using Python Pytest libraries. Developed lake house using Azure Databricks, Azure Data Lake and Delta Lake.