Requires two 2 years of experience with Large Scale Distributed Data Processing; Data Processing using Scala or Spark for Grouping and Aggregations; Data processing pipelines; Setting up scalable Spark clusters; Stream processing pipelines; Scala; Developing REST APIs; Java; Apache Spark; Sybase and DB2; SQL; Hive; Oozie; Kafka; Gradle; Github; K8S Kubernetes; Jenkins; SOLR; HDFS; Docker; Openstack; Ubuntu and CentOS; Linux; Mongo DB; Jupyter; Unit and Integration Testing; and Ansible.
Categories: eb3
0 Comments