Prashant Kumar
Big Data Engineer @John Deere | Spark | SQL | PySpark | Hadoop | Hive
- Role
- Senior Engineer Ii - Dsea at John Deere
- Location
- Pune District, MH, IN
- LinkedIn followers
- 500 followers
About Prashant Kumar
Passionate about solving real-world problems through data, I’m a Data Engineer with 3…
Experience
Senior Engineer Ii - Dsea
Nov 2023 — Present · Pune, IN
Migrated 500+ GB of on-prem CSV data to AWS S3, building scalable ingestion pipelines using Databricks and PySpark. Enabled centralized access to previously local data and accelerated analytics by moving dashboards to cloud, reducing processing time by 95%.• Developed 10+ Tableau dashboards on Databricks pipelines, processing over 50 billion records from data lake to enhance data visibility for 20+ stakeholders and driving 15%+ improvement in design efficiency and reduced field failures.• Conducted RCA of bad data, querying approximately 200M records, and successfully implemented a metadata driven Data Quality Management tool using Databricks. This initiative significantly reduced data issue resolution time by 75% improving overall data quality for the product.• Identified discrepancies in engine size by cross-referencing 2 data sources, optimized engine performance assessments by ensuring alignment between various data sources, ultimately contributing to 25% more operational efficiency through actionable insights.• Streamlined testing and troubleshooting of 50+ test machines, reducing decision-making time for improvements by 30%, enhancing system reliability, and minimizing post-release issues.• Engineered a Python and Spark-based pipeline on a local system to collect and process 1M+ diagnostic records weekly from cloud sources, generate over 50.mdf and.csv files by machine and week, store them in customer-specified directories, & automate the archival of historical files weekly for streamlined data management.• Developed a success and failure tracking system, reducing troubleshooting time by 40% and strengthening debugging efficiency across 20+ R scripts which updates datasets for Tableau dashboards.• Consolidated 50 Million+ rows of unstructured data into tabular formats using Databricks and designed 10+ Tableau visualization to track 15+ KPIs, improving monitoring efficiency by 25% for the agricultural and construction vehicle configuration team.
Education
TECHNO INDIA UNIVERSITY
Bachelor of Technology - BTech, Computer Science
2016 — 2020
National Institute of Technology Hamirpur
Master of Technology - MTech, Computer Science
2020 — 2022
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.