Experience in the following representing terabyte scale raw data as features and labels; using machine learning libraries to generate models; predicting outputs by applying models on unlabeled terabyte scale data; designing terabyte scalable real time and batch processing application with Spark; developing scalable data pipelines to load, transform, store and enrich inhouse big datasets on cloud platforms; design, build, and maintain scalable machine learning pipelines that can continuously operate terabyte datasets.
Categories: eb3
0 Comments