Saumya Maurya

Data Engineer @Bliss Ladders

Barrie, ON, CA
MOBILE NUMBERS
+91 *********19

Signup · Get unlimited contacts

WORK HISTORY

Jun 2024 — Present

Data Engineer @Bliss Ladders

View department →

ON, CA

Building and optimizing cloud-native data infrastructure on AWS from raw ingestion to executive-facing analytics. • Architected a multi-source AWS lakehouse (CSV, Parquet, APIs) consolidating 6+ upstream systems, increasing BI self-service access by 3×. • Re-engineered legacy ETL to serverless AWS Glue + Apache Spark + Redshift Spectrum, boosting processing throughput by 55% and enabling real-time analytics. • Optimized PySpark transformation scripts via parallel processing and memory tuning, cutting runtime by 40% on 500M+ record datasets. • Reduced dashboard query latency by 60% through Redshift partitioning, indexing, and distribution key redesign, improving executive reporting SLAs. • Implemented AWS Glue Data Catalog + Lambda automation for end-to-end data lineage and schema evolution, achieving 100% governance compliance across all production environments.

EDUCATION

N/A

Georgian College

𝐏𝐨𝐬𝐭𝐠𝐫𝐚𝐝𝐮𝐚𝐭𝐞 𝐃𝐞𝐠𝐫𝐞𝐞 𝐢𝐧 𝐀𝐫𝐭𝐢𝐟𝐢𝐜𝐢𝐚𝐥 𝐈𝐧𝐭𝐞𝐥𝐥𝐢𝐠𝐞𝐧𝐜𝐞, Artificial Intelligence (AI)

N/A

Georgian College

Postgraduate Degree in Mobile Application Development, Mobile Application Development

2018 — 2021

The Maharaja Sayajirao University of Baroda

𝐁𝐚𝐜𝐡𝐞𝐥𝐨𝐫'𝐬 𝐨𝐟 𝐂𝐨𝐦𝐩𝐮𝐭𝐞𝐫 𝐀𝐩𝐩𝐥𝐢𝐜𝐚𝐭𝐢𝐨𝐧𝐬 (𝐁𝐂𝐀), Computer Programming/Programmer, General

ABOUT SAUMYA MAURYA

I build data pipelines that perform, not just pipelines that run.Over 3+ years across Azure, AWS, and GCP, I\'ve architected ETL/ELT systems that delivered 40–60% performance improvements, handled 500M–5B+ record datasets, and met production-grade governance standards. My stack spans Azure Data Factory, Synapse Analytics, AWS Glue, Redshift, Kafka, and PySpark with a strong focus on reliability, lineage, and scale.A few things I\'ve shipped: • Re-engineered legacy ETL to serverless AWS Glue + Spark pipelines → 55% throughput gain • Reduced dashboard query latency by 60% via Redshift partitioning and distribution key redesign • Implemented Azure Purview for end-to-end metadata lineage → 90% traceability coverage • Built real-time IoT streaming pipelines with sub-second anomaly detectionI hold a 4.0 GPA postgraduate diploma in AI from Georgian College and AWS certifications, and I bring ML pipeline awareness to every data architecture decision.Currently on an Open Work Permit and actively exploring Data Engineer and Data Analyst roles across Canada (remote-friendly or on-site).Let\'s build something solid.

This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.