G Nagalaxmi
Data Engineer @Cleveland Clinic
Signup · Get unlimited contacts
WORK HISTORY
Data Engineer @Cleveland Clinic
I developed end-to-end data engineering solutions to process large-scale healthcare datasets while ensuring compliance with HIPAA regulations. I designed and optimized ETL pipelines using Spark, Hive, and HDFS, and integrated hybrid data solutions across Azure (ADF, HDInsight, Blob Storage) and AWS (S3, Glue, Lambda). I implemented real-time streaming with Kafka and Spark Streaming for patient monitoring data, automated workflows using Oozie and Shell scripts, and supported data governance with Collibra and Apache Atlas. My contributions empowered clinical teams with near real-time dashboards (Power BI, Tableau) for patient health, treatment outcomes, and operational performance.
EDUCATION
Southeast Missouri State University
Master's degree
ABOUT G NAGALAXMI
I am a Data Engineer with around 7 years of experience designing and optimizing large-scale data pipelines, ETL workflows, and cloud-based solutions to drive advanced analytics and business intelligence. I specialize in building robust, scalable, and high-performance data platforms using Big Data technologies (Hadoop, Spark, Hive, Kafka, HBase) and cloud services (AWS, Azure, GCP). My expertise spans across data modeling, warehousing, real-time streaming, and workflow orchestration with tools like Apache Airflow, Oozie, and AWS Step Functions. With hands-on experience across industries including finance, healthcare, and e-commerce, I have successfully delivered enterprise-grade data solutions that power decision-making, real-time analytics, and compliance-driven reporting.
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.