Binish Abraham
Big Data | Data Engineer | Data Science
- Role
- Big Data Developer Architect at Bank of America
- Location
- Queens, NY, US
- LinkedIn followers
- 500 followers
About Binish Abraham
A focused and detail oriented Big Data Engineering professional. Focused experience in distributed processing and cloud technologies. Experience in various ML algorithms and statistical models. 13+ Years of total experience in vast technology space and consulting roles for clients from different domains.MS in Data Science from St. John’s University, New YorkEngineering degree in Computer Science and EngineeringExperience in design and development of applications using distributed technologies like Hadoop, HBase, Hive, Cassandra, Storm, Spark, Neo4j, MongoDB, ImpalaExperience in architecting and developing distributed solutions using grid technologies and public cloud like AWS and GCPWorked for clients from various domains in the US - Bank of America, Turner, BJs, Comcast Corp, Siemens Healthcare, New York Times, ITHAKA, Mattel, Bank of New York MellonExperience in search technologies like Elasticsearch and SOLR Excellent exposure to Linux OS internals and RHCE certifiedExperience in Docker Container technology and KubernetesExperience in machine learning algorithms KNN, K-Means, Naïve Bayes, LDA, CNN, transfer learning, web data mining and NLP using Gensim LDAFamiliar with t-SNE, DNN and using libraries like TensorFlow, Scikit-learnExperience in Databricks, Jupyter Notebooks, Google Colab and Anaconda, Experience in Statistical analysisProgramming languages - Python, Java, R, ScalaFamiliar with PCA, SVD, SVM, Spectral clustering
Experience
Big Data Developer Architect
Jun 2022 — Present · New York, NY, US
Designed and developed hive and Impala queries, Python scripts, Scala code changes, Shell scripts for ETL purpose based on end user requirements• Designed, developed and modified SQL queries based on HQL (Hive Query Language) for business needs which executes in spark environment• Architected and implemented a robust Hadoop infrastructure, deploying Hadoop Distributed File System (HDFS) and setting up clusters for efficient data storage and processing• Implemented robust data quality controls, including validation checks and error handling mechanisms, to ensure data accuracy and consistency throughout the platform• Worked on providing data for models by running multiple data pipelines which extract data from multiple data sources and export the required data and combine them• Worked on data transformations and loading using PySpark and Python for ML projects• Analyzed Spark and Hadoop logs using UI and SSH nodes to find issues and resolve them• Maintained proper code using Bitbucket and participated in monthly release activities
Education
Ilahia College of Engg Cochin
B Tech, Computer Science
2003 — 2006
St. John's University
Master of Science - MS, Data Science
Govt Polytechnic College
Diploma, Computer Science
1999 — 2002
Skills
- Cassandra
- Nosql
- Rhce
- Hbase
- Shell Scripting
- Hadoop
- Amazon Web Services (Aws)
- Grid Computing
- Hive
- Amazon Web Services
- Cloud Computing
- Riak
- Java
Find verified contacts for anyone on LinkedIn
Unifers gives sales teams verified emails and direct dials, enriched profiles, and outreach that lands in the inbox.
Free plan included · No credit card required
This profile is compiled from publicly available professional sources. Unifers is not affiliated with or endorsed by LinkedIn. Request removal of this profile.