Ashok Kumar P
Data Scientist
- Role
- Data Scientist at Capg
- Location
- Charlotte, NC, US
- LinkedIn followers
- 500 followers
About Ashok Kumar P
Overall, 14+ years of experience in IT industry including 10+ years of exp As data engineer using Microsoft Azure with Databricks, Python, Spark, PySpark, Spark SQL, Databricks workspace for Business Analytics, Manage Clusters in Databricks, Managing the Machine Learning Lifecycle and 4 + years of exp as data engineer using Google Cloud Platform with services Big Query, Data flow, Data Fusion, Cloud storage, Cloud Functions, Pub/Sub, Data Proc, Apache Airflow, Apache Beam, Big Table and Kubernetes. • Hands on exp Data extraction (extract, Schemas, corrupt record handling and parallelize code), transformations and loads (user -defined functions, join optimizations) and Production (Optimize and automate Extract, Transform and Load) • Hands on exp on Unified Data Analytics with Databricks, Databricks Workspace User Interface, Managing Databricks Notebooks, AWS S3, AWS Glue, Delta Lake with Python, Neo4j, Delta Lake with Spark SQL.• Worked on projects in waterfall and agile methodology.• Exp on Migrating SQL database to Azure data Lake, Azure data lake Analytics, Azure SQL database.• Exp in Developing Spark applications using Spark – SQL in Databricks for data• Good understanding of Bigdata Hadoop and Yarn architecture along with various Hadoop Demons such as Job Tracker, Task Tracker, Name Node, Data Node, Resource/Cluster Manager and Kafka.• Experience in designing and implementing data engineering architectures using GCP (Google Cloud Platform),• Such as Google Big Query and Cloud Storage.• Expertise migrating Hadoop data from on-premises to GCP provides additional context for incrementally moving your data to google cloud (GCP).• Expertise migrating jobs from on premises to Data proc in GCP• Expertise migrating data from Hbase to GCP Bigtable• Expertise with large datasets and solving difficult analytical problems is a primary task of a Google data engineer.• Expertise in understanding, solving big data problems using Hadoop ecosystem components such as HDFS, Map Reduce, Hive, Oozie, Autosys.• Extensive Knowledge on developing Spark Streaming jobs by developing RDD’s (Resilient Distributed• Datasets) using Scala, PySpark and Spark-Shell.• Extensively worked with GCP data services and tools such as Google Cloud Storage, BigQuery, Dataflow,• Dataproc, and Cloud SQL.• Designed and Developed Spark work on Aws using Scala for data pull from AWS S3 bucket and• Snowflake applying transformations on it.• Rich experience in Banking, Retail, E-commerce, Telecom, Health Insurance
Experience
Data Scientist
Jun 2023 — Present · US
Education
JNTU Anantapur
Bachelor of Engineering - BE, Electrical, Electronics and Communications Engineering
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.