Gowtham Mukkara
Data Engineer at CITI | Transforming Data into Actionable Insights | PySpark | SQL | Python | AWS | Azure Databricks | Palantir Foundry | PowerBI | Tableau |
- Role
- Data Engineer at Tata Consultancy Services
- Location
- Dallas, TX, US
- LinkedIn followers
- 500 followers
Experience
Data Engineer
Jun 2024 — Present · Dallas, TX, US
Curated, cleansed, and transformed large-scale datasets using SQL, Python, Azure Databricks, and Azure Data Factory, improving data integrity by 25% and enhancing analytics accuracy.Developed a Python framework to automate schema file generation for data modeling using Model-Driven Development (MDD), reducing sprint effort by 1.5 days (20%) and improving efficiency in the data engineering workflow.Optimized SQL performance by implementing indexing, partitioning, query refactoring, and denormalization, reducing execution time and improving system efficiency.Developed and orchestrated ETL pipelines in Azure Data Factory, automating data ingestion and transformation, reducing manual configurations by 20%, and ensuring seamless data flow.Designed and optimized Azure-based ETL workflows using Azure Data Factory (ADF) and Databricks, implementing parallel processing and dynamic data partitioning to improve pipeline execution speed and scalability.Automated metadata-driven pipeline orchestration in Azure Data Factory, enabling dynamic mapping of data sources, reducing manual intervention, and improving reusability across multiple workflows.Designed and implemented CI/CD workflows using PyCharm, Bitbucket, and Azure DevOps, streamlining data pipeline deployment and ensuring efficient version control.Leveraged Apache Spark in Azure Databricks to enhance data processing for both batch and streaming pipelines, optimizing computational performance.Integrated robust logging, monitoring, and alerting mechanisms, improving failure detection and minimizing downtime in critical data pipelines.Executed unit testing and validation using PyTest and Databricks Notebooks, reducing post-deployment defects by 30% and improving data reliability.Automated schema evolution and data validation to handle structural changes dynamically, preventing ingestion failures and ensuring pipeline adaptability. Collaborated in Agile environments, actively engaging in Sprint Planning.
Education
George Mason University
Master's Degree
REVA University
Bachelor's degree, Computer Science
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.