Revanth M
Big Data Engineer at Core (Looking for remote)
- Role
- Sr Data Engineer at Core Scientific
- Location
- Palo Alto, CA, US
- LinkedIn followers
- 500 followers
About Revanth M
Around 7+ years of total IT Experience, including 5+ years of extensive experience in Building and executing analytics and reporting across platforms to identify user behavior and analyze trends, patterns, and shifts in user behavior, both independently and in collaboration with product managers and data analytics resources.Developed experimental data models/designs to help answer unforeseen questions that will influence decision-making in a rapidly changing business environment.Expertise in all Hadoop Ecosystem components- Hive, Hue, Pig, Sqoop, Impala, Flume, Zookeeper, Oozie, Airflow, and Apache Spark. Good expertise in Creating, Debugging, Scheduling, and Monitoring jobs using Airflow and Oozie.Excellent understanding/knowledge of Hadoop architecture and various components such as HDFS, Job Tracker, Task Tracker, Name Node, Data Node, YARN, MapReduce programming paradigm, etc. Experience in handling large datasets using Partitions, Spark in memory capabilities, Broadcasts in Spark with Scala and python, Effective and efficient Joins, Transformations and other during ingestion process itself.Experience in developing data pipelines using Pig, Sqoop, and Flume to extract the data from weblogs and store in HDFS and accomplished developing Pig Latin Scripts and using HiveQL for data analytics.Experience in importing and exporting data from different RDBMS like MySQL, Oracle, and SQL Server into HDFS and Hive using Sqoop.Experience in developing custom MapReduce programs using Apache Hadoop to perform Data Transformation and analysis as per requirement.Hands-on experience on Scala language features - Language fundamentals, Classes, Objects, Traits, Collections, Case Classes, High Order Functions, Pattern Matching, Extractors etc.Experience in creating PIG and HIVE UDFs using Java in order to analyze data sets.Experience in Spark Streaming in order to ingest real-time data from multiple data sources into HDFS.Experience in Design & Development, tuning, and maintenance of NoSQL databases such as MongoDB, HBase, Cassandra, and its Integration with Hive.Worked on reading multiple data formats on HDFS using Spark API.Design, develop, deploy, and manage a reliable and scalable data analysis pipeline, using technologies including AWS, open-source tooling, and custom-built frameworks.Excellent experience in the Application Development and Maintenance of SDLC projects using various technologies such as Python, Java/J2EE, Scala, JavaScript, Data Structures, UNIX shell scripting.
Experience
Sr Data Engineer
Oct 2022 — Present
Unified data analytics with Databricks, Databricks workspace user interface, Managing Databricks notebooks, Delta Lake with Python, Spark, and SQLWorking on spark architecture with Databricks, Structured streaming with Delta Live Tables and Configuration, Managing Cluster in DatabricksResponsible for creating on-demand tables on cloud-based S3 files using Lambda Functions and AWS Glue using Python.Data Extraction (extract, Schemas, Corrupt record handling and Ingestion automation), Transformation and loads (user defined function, join optimization) and production (Optimize and auto extract, transform and load)Developed a data platform from scratch and took part in requirement gathering and analysis phase of the project in documenting the business requirements.Worked on visualization dashboards using power BI (DAX) and Tableau
Education
Northwestern Polytechnic University
Master of Science - MS, Electrical and Electronics Engineering
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.