Requires a Bachelors degree or Masters degree in Computer Science, Information Technology or related field or foreign equivalent. Requires five 5 years with Bachelors or three 3 years with Masters providing data engineering on the Data Lake or Hadoop using Spark, Impala and HDFS Hadoop file system; building data pipelines to load and manipulate data onto the Data Lake; coding data pipelines using SQL or Impala, Java, and Scala or Spark; building derivative data structures onto Hadoop; designing quality control tests; designing partitioning and bucketing strategies; and with Bachelors 3 years, with Masters 1 year analytics with data on Hadoop, including for data science, machine learning and statistical use cases. Work MF 8.30 am 5pm 37.5 hours week.

Categories: eb3

0 Comments

Leave a Reply

Avatar placeholder

Your email address will not be published. Required fields are marked *