Saurav Roy
Team Lead (Data) at GlobalLogic
- Role
- Technical Lead (Data Engineering) at Tata Consultancy Services
- Location
- Gurugram, HR, IN
- LinkedIn followers
- 500 followers
About Saurav Roy
As a Technical Lead, my current work revolves around performing data engineering ETL and ELT pipelines. My primary technical stack involves Python, PySpark, SQL, AWS, Snowflake, Airflow and DBT. “Data is the new oil.” — Clive Humby
Experience
Technical Lead (Data Engineering)
Jul 2023 — Present · Gurugram, IN
Building data transformation and ELT pipelines using Snowflake and DBT models to create real-time, event-driven data ingestion pipelines.2. Designed and optimised scalable data models and warehouses in Snowflake to support advanced analytics.3. Designed DAGs in airflow to automate and optimise end-to-end data pipeline execution.4. Designed and developed ETL pipelines using PySpark, Python, SQL, AWS, Apache hive, and Exadata tables for dual domains i.e. Data analytics and Service Improvements.5. Facilitated multiple ETL pipelines handling more than 20 million records per run within SLA including storage in AWS S3 and loading data to hive tables.6. Implemented Python scripts that automated routine data transformation tasks, saving over 200 hours of manual effort annually.7. Resolved critical issues in production systems reducing downtime by 40% through problem solving.8. Led a team of 4-5 members to deliver multiple projects of medium to high impact including managing project-based timelines and deliverables using Jira stories, confluence and storyboards.9. Collaborated closely with cross-functional teams to understand data requirements resulting in delivery of nearly 100% uptime for critical analytics services.10. Designed and developed data processing pipelines to perform historical loads of the previous 10 years of data in datalake using SQL, PySpark and AWS S3.11. Spearheaded and successfully migrated more than jobs from Python 3.7 to Python 3.9 including functional and non-functional testing.12. Developed the ETL pipelines for performing data archival consisting of data from both S3 and hive tables to improve the performance of data load speeds by 80%.13. Developed and designed multiple AdHoc service improvement tasks improving their performance by more than 70% to 90% using PySpark and other elements like multi-threading and parallel processing.
Education
Apeejay Stya University
Bachelor of Technology (B.Tech.), Electrical, Electronics and Communications Engineering
2013 — 2017
Skills
- Css
- Illustrator
- Script Writing
- Matlab
- Research and Development (R&D)
- Film Editing
- Microsoft Office
- Satellite Robotics
- C
- Atmel Avr
- Html5
- Creative Writing
- Jquery
- Screenwriting
- Acrylic Painting
- Microsoft Powerpoint
- Radio Production
- Gui Development
- Javascript
- Lightroom
- Microsoft Word
- Adobe Photoshop
- Sound Editing
- Photoshop
- Html
- Arduino
- Serial Communications
- Mobile Robotics
- Video Editing
- Embedded C
- Text Editing
- Research
- Tv Production
- Powerpoint
- Robotics
- Microsoft Excel
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.