Madhu Vemula

Senior Data Engineer Analyst @TriNet

Chandler, AZ, US
MOBILE NUMBERS
+91 *********19

Signup · Get unlimited contacts

WORK HISTORY

Nov 2022 — Present

Senior Data Engineer Analyst @TriNet

View department →

Phoenix, AZ, US

Led end-to-end cloud data engineering processes, encompassing requirements gathering, system analysis, design, development, testing, and deployment on Google Cloud Platform (GCP). Established robust data pipelines utilizing Google Cloud Dataflow services to transfer data from on-premises sources to Google Cloud SQL for seamless data orchestration. Executed the construction of data pipelines employing GCP services like Dataflow to migrate data from legacy SQL servers to Google BigQuery using Dataflow and Cloud Dataprep. Constructed intricate ETL jobs for visual data transformations through data flows, leveraging Google Cloud Dataprep and Google Cloud BigQuery. Leveraged a variety of activities in Google Cloud Dataflow for data movement, transformation, and control purposes, including Dataflow, Data Fusion, Apache Beam, and Cloud Composer. Developed pipelines to extract and load data from on-premises systems to Google Cloud Storage. Implemented comprehensive data copy operations with a focus on hierarchical transformations and error handling within Google Cloud Dataflow.Employed GCP Dataflow and Cloud Dataprep for executing data copy activities and transformations within the data pipeline. Proficient in Google Cloud Dataflow activities such as Lookup, Conditional operations, Looping constructs, Variable management, Metadata retrieval, Filtering, and Wait operations. Applied Python for data analysis, statistical modeling, and hypothesis testing to derive actionable insights from diverse datasets.Extensively used Google Cloud Key Management Service (KMS) to configure linked service connections securely. Configured and managed Google Cloud Dataflow Triggers and scheduled pipelines for automated data workflows. Monitored scheduled GCP Dataflow pipelines and set up alerts for failure notifications. Contributed to the design and development of real-time data processing using Google Cloud Stream Analytics.

ABOUT MADHU VEMULA

As a senior data engineer l, I am responsible for designing, developing, testing, and deploying end-to-end data pipelines using various cloud services and tools. I have over 8 years of experience in data engineering, working with different platforms, sources, and formats of data. My core competencies include PySpark, SQL, AWS, Azure, and ETL development.In my previous roles, I have created data pipelines that ingest, curate, and provision data from MySQL, Oracle, MongoDB, SFTP, and legacy systems, using services like S3, EC2, EMR, Redshift, Athena, Glue, DynamoDB, RDS, IAM, Data Factory, Databricks, and SQL Server. I have also developed PySpark applications, Tabular Model infrastructure, and Power BI dashboards to perform data transformations and visualizations. I am passionate about solving complex data problems and delivering value to federal agencies and clients. I bring diverse perspectives and experiences to the team, as I have worked in various industries, such as finance, health and safety, and IT. I enjoy collaborating with business users, IT teams, and ETL teams to understand their requirements and provide technical solutions.

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.