Position: Data Engineer Company: Petrel AI Location: Almaty Salary: 1-1.5M KZT
Role The firm is looking for highly professional software developers, who has relevant experience working with real-time data ingestion, building fault tolerant distributed systems for large datasets. Previous work experience in industrial company with IoT would be a plus for a candidate.
Skills Required • Proficient understanding of distributed computing principles • Management of Hadoop cluster, with all included services • Ability to solve any ongoing issues with operating Hadoop clusters • Proficiency with Hadoop 2.x, MapReduce, HDFS • Experience with building stream-processing systems, using solutions such as Storm or Spark-Streaming • Good knowledge of Big Data querying tools, such as Pig, Hive and Impala • Experience with Spark, Jenkins, Sqoop, HDInsight, U-SQL, Java, Python, JSON, Linux • Experience with integration of data from multiple data sources • Experience with NoSQL databases • Knowledge of various ETL techniques and frameworks Extract Transform Load • Experience with various messaging systems Apache Kafka • Experience with Big Data ML toolkits • Good understanding of Lambda Architecture, along with its advantages and drawbacks • Advanced data analysis skills including advanced SQL query capabilities • Deep experience in data modeling, data analysis and relational database design • Demonstrated ability to identify business and technical impacts of user requirements and incorporate them into the project schedule • Ability to work both independently and as part of a team • Ability to work under pressure and to independently handle multiple projects and deadlines • Experience working with large datasets
Please feel free to contact me in case of your interest or if you can suggest someone for this role.