Job Details
Ideally, you should have:
- 2-6 years of overall industry experience
- Minimum of 2 years of experience building and deploying large scale data processing pipelines in a production environment
- Strong domain modeling and coding experience in Java /Scala / Python.
- Experience building data pipelines and data-centric applications using distributed storage platforms like HDFS, S3, NoSql databases (Hbase, Cassandra, etc) and distributed processing platforms like Hadoop, Spark, Hive, Oozie, Airflow, Kafka, etc in a production setting
- Hands on experience in (at least one or more) MapR, Cloudera, Hortonworks and/or Cloud (AWS EMR, Azure HDInsights, Qubole, etc.)
- Knowledge of software best practices like Test-Driven Development (TDD) and Continuous Integration (CI), Agile development
- Strong communication skills with the ability to work in a consulting environment is essential
Reference: www.eduinq.com
To get automatically referred, kindly apply with www.eduinq.com and submit your resume once a day.