Jennifer Lawrence
Hadoop Developer at Infosys
- Role
- Big Data Developer at Infosys
- Location
- Houston, TX, US
- LinkedIn followers
- 500 followers
About Jennifer Lawrence
I am IT Professional with 6.9 years of experience in Big Data and Hadoop, Java and Scala/Spark development. I work on big data projects, where my responsibilities usually include solution and data architecture, integration, streaming, real time and batch processing.• Developed solutions to process data into HDFS process within Hadoop and emit the summary results from Hadoop to downstream analytical systems• Experience in BigData, Spark, Hadoop, MapReduce, Sqoop, Flume, Hive, Oozie, MongoDB, AeroSpike• Experience in Cloud Application development, Cloud automation and cloud deployment • Experience in Big data analysis using Pig and User defined functions (UDF)• Experience in working with different data sources like Flat files, XML files, Log files and Databases• Experience in using the AWS, Google Cloud and platforms• Explored various concepts and tools in the field of real time data analytics using Apache Storm• Having experience on using OOZIE to define and schedule the jobs• Experience in Data structures, Design patterns• Experience in using GIT, Maven, Gradle, SBT building toolsSpecialties• Hadoop, Storm, Spark, HDP, CDH• MPP and SQL databases, data architecture, integration, DWH, BI, advanced analytics• Apache Spark : Python API, Scala API, DataFrames, Spark SQL, MlLib• Big data architecture, implementation, delivery• Hadoop, Hive, Sqoop, Zookeeper, Flume • Delivery management, DevOps• Distributions: Cloudera/Hortonworks• AWS• Programming Languages : Java / Python / Data Structures / SQL / Scala• Scripting Language - Shell Scripting• Database Technologies - MongoDB/NoSQL/Oracle / MySQL
Experience
Big Data Developer
Feb 2012 — Present · Houston, TX, US
Shared responsibility for administration of Hadoop, Hive and Pig.• Managed and reviewed Hadoop log files.• Provided design recommendations and thought leadership to sponsors/stakeholders that improved review processes and resolved technical problems.• Exported the analyzed data to the relational databases using Sqoop for visualization and to generate reports by our BI team.• Tested raw data and executed performance scripts.• Wrote PIG scripts and executed by using Grunt shell.• Analyzed the web log data using the HiveQL to extract number of unique visitors per day, page views, visit duration, most purchased product on website.• Developed PIG UDFs for the needed functionality and custom Pigsloader known as timestamp loader. • Exported the analyzed data to the relational databases using Sqoop for visualization and to generate reports by our BI team.• Performed ETL processes using Pentaho.• Responsible to manage data coming from different sources. • Involved in Installing and configuring Kerberos for the authentication of users and Hadoop daemons.• Integrated Oozie with the rest of the Hadoop stack supporting several types of Hadoop jobs out of the box (such as Map-Reduce, Pig, Hive, and Sqoop) as well as system specific jobs (such as Java programs and shell scripts).
Education
Texas Tech University
Master of Computer Applications (MCA)
2007 — 2010
Skills
- Apache
- Hadoop
- Apache Spark
- Apache Pig
- Mongodb
- Big Data Developer
- Large-Scale Data Analysis
- Apache Storm
- Big Data
- Data Science
- Scala
- Big Data Analytics
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.