Senior Data Engineer (Hadoop, Spark) @ GFT Poland
Core
Designing, developing, and maintaining scalable data solutions and pipelines within a dynamic DevOps environment.
Role type
Senior Data Engineer (Big Data)
Builds
Scalable data pipelines and solutions
Domain
Big Data / Data Engineering
Deliverable
production ML models | infrastructure
Required skills
PySpark, Scala, Apache Airflow, Hadoop ecosystem (Spark, Hive, YARN), ETL frameworks, SQL, RESTful services, Unix/Linux operations, Git, Jenkins, Ansible, Big data modelling, Spark streaming, Java APIs, PL/SQL
Preferred skills
Elasticsearch, Cloud design patterns, DevOps practices, Agile methodologies (Scrum, Kanban)
Technologies
PySpark, Scala, Airflow, Hadoop, Spark, Hive, YARN, Elasticsearch, Java, Python, PL/SQL, Linux, Unix, Git, GitHub, Jenkins, Ansible, JIRA
Responsibilities
Design and develop software using PySpark, Automate testing of components, Conduct code reviews and mentoring, Provide production support and troubleshooting, Implement tools for performance and monitoring, Collaborate with Business Analysts on requirements
Seniority
Senior, hands-on IC