Scala development and design using Scala 2.10+ or Java development and design using Java 1.8+
Experience with most of the following technologies (Apache Hadoop, Scala, Apache Spark, PySpark, Spark streaming, YARN, Kafka, Hive, Python, ETL frameworks, Map Reduce, SQL, RESTful services).
Sound knowledge of working Unix/Linux Platform.
Hands-on experience building data pipelines using Hadoop components - Hive, Spark, Spark SQL, PySpark.
Experience with industry-standard version control tools (Git, GitHub), automated deployment tools (Ansible & Jenkins), and requirement management in JIRA.
Understanding of big data modeling techniques using relational and non-relational techniques.
Experience in Debugging the Code issues and then publishing the highlighted differences to the development team/Architects.