Experience in Python data related programming using various packages in python including ggplot2, caret NLP, plyr, pandas, numpy, seaborn, scipy, matplotlib, scikitlearn, Beautiful Soup; Proficient with Big Data Analysis, mapping source and target systems for data migration efforts and resolving issues relating to data migration using Hive and Map reduce, SQL; Strong data architecting skills designing datacentric solutions and hands on experience with big data tools like Hadoop, Snowflake, Pandas, Spark, Hive, Pig, Impala, Pyspark, SparkSql; Good knowledge and experience of cloud computing and storage systems, AWS EC2, Sagemaker, Redshift, S3 and EMR; Expertise in managing entire data science project life cycle and actively involved in all the phases of project life cycle including data acquisition, data cleaning, data migration, data mapping, data engineering, data validation, features scaling, features engineering, statistical modeling decision trees, regression models, neural networks, SVM, clustering, dimensionality reduction using Principal Component Analysis and Factor Analysis, testing and validation using ROC plot, Kfold cross validation and data visualization.

Categories: eb3

0 Comments

Leave a Reply

Avatar placeholder

Your email address will not be published. Required fields are marked *