Big Data Engineer
Core
Design, build, and maintain large-scale data processing systems and pipelines.
Role type
Big Data Engineer
Builds
Scalable data pipelines and ETL processes
Domain
Big Data / Distributed Systems
Deliverable
production ML models
Required skills
Python, Java, Scala, Apache Spark, Hadoop, Kafka, Flink, SQL, relational databases, NoSQL databases, ETL processes, data warehousing, cloud platforms (AWS, Azure, GCP), distributed systems architecture
Preferred skills
Airflow, data lake architectures, Docker, Kubernetes, real-time data processing
Technologies
Hadoop, Spark, Kafka, Flink, Airflow, Docker, Kubernetes, AWS, Azure, GCP
Responsibilities
Design and develop scalable data pipelines and ETL processes; Work with large, complex datasets in distributed environments; Implement data processing solutions using big data tools; Optimize data workflows for performance, scalability, and reliability; Integrate data from multiple sources including APIs and databases; Monitor and troubleshoot data pipeline issues
Seniority
Mid-level, hands-on IC