Data Engineer (25396)
Posted yesterday · 0 applicants
Saving, applying or scoring takes a few seconds to set up your free account.
The role in plain words
- Relevant academic background
- At least 4 years of hands-on experience as a Data Engineer or in a Big Data development role
- Cloudera platform
- Apache Spark
- Apache Hive
- Apache Airflow
- Apache Kafka
- Data infrastructure preparation for Machine Learning and AI projects
- Google Cloud Platform (GCP)
Extracted from the job description · kept up to date automatically
Who this suits
Full job description
Original listing · kept for referenceWho We Are:
Yael Group is a leading group of companies in the Israeli market, providing advanced technological solutions across a wide range of domains to organizations in all sectors.
Job Description:
Join the Data & AI unit in a key role responsible for building a large-scale enterprise Data Lake from end to end in an on-premises environment.
• Design, develop, and maintain scalable data infrastructures based on the Cloudera platform.
• Design and implement enterprise data architecture, including high-volume data ingestion, processing, and Data Pipelines in a distributed environment.
• Develop and maintain data processing pipelines to support advanced analytics capabilities and AI model development.
• Work in a cutting-edge technology environment combining Big Data, Data Engineering, and Artificial Intelligence.
• Collaborate closely with Data, Analytics, and Development teams to build enterprise-grade data infrastructure.
• Hybrid work model.
Job Requirements:
• Relevant academic background.
• At least 4 years of hands-on experience as a Data Engineer or in a Big Data development role.
• Proven experience working with the Apache ecosystem, including core technologies such as Apache Spark, Apache Hive, and Apache Impala.
• Strong proficiency in developing ETL/ELT data pipelines using at least one of the following programming languages: Python, Scala, or Java.
• Expertise in writing complex analytical SQL queries and optimizing performance in distributed data environments.
• Deep understanding of distributed systems architecture and data modeling.
• Experience with workflow orchestration tools such as Apache Airflow – an advantage.
• Hands-on experience with streaming data processing using Apache Kafka – an advantage.
• Experience preparing data infrastructure for Machine Learning and AI projects – an advantage.
• Experience working in hybrid environments and familiarity with Google Cloud Platform (GCP) services – an advantage.
• Willingness and availability to support critical systems when required, including after-hours and weekend work.
Questions about this role
- This listing did not state a salary. We only show pay when the employer publishes it.
- Relevant academic background, At least 4 years of hands-on experience as a Data Engineer or in a Big Data development role, Cloudera platform, Apache Spark, Apache Hive