Hemanth Reddy
Hadoop Lead Developer Associate Data Scientist @Tufts Health Plan
Signup · Get unlimited contacts
WORK HISTORY
Hadoop Lead Developer Associate Data Scientist @Tufts Health Plan
CA
Worked on multiple projects using various Big Data Technologies.• Worked on Data Scientist activities and developed different scatter graphs using R-Studio.• Worked on various types of machine learning algorithms.• Created automated python scripts to validate the data flow through elastic search.• Experience in AWS cloud environment on S3 storage and EC2 instances. • Worked on evaluation and analysis of Hadoop cluster and different big data analytic tools including Pig, HBase database and Sqoop.• Created HBase tables to store various data formats of data coming from different portfolios.• Experience in managing and reviewing Hadoop log files.• Setting up the ELK (ElatsticSearch, Logstash, Kibana) Cluster.• Django Framework used in developing web applications to implement the MVC architecture • Design and Development of adapters to inject and eject data from various data source to/from Kafka. • Design and development of HBase tables according to various needs of the tenants while taking into consideration various issues related to performance. • Interact with various business teams to document the requirements for HBase tables• Developed Spark applications to move data into HBase tables from various sources like Relational Database or Hive• Maintaining and scheduling various Spark, MapReduce and sqoop jobs according to business needs and maintaining data consistency using oozie• Optimize the Spark applications both while developing and while submitting it to the cluster with various environment arguments• Developed and written ApachePIG scripts and HIVE scripts to Load, Store and Process the Data• Worked closely with analysts and architects to understand the business requirements for system enhancements and Data Analytics (Using HadoopOLAP principles)• Working knowledge in writing Pig\'s Load and Store functions• Developed sqoop scripts that can store the data from relation to hive and hbase directly• Demo various functional capabilities of Spark
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.