Cloudera Developer
Core
Design, develop, and implement data solutions using Cloudera technologies such as Hadoop, Spark, and Hive to optimize data pipelines and ensure data quality.
Role type
Senior Cloudera Developer (Data Engineer)
Builds
Data pipelines, data processing workflows, and data solutions on Hadoop/Spark/Hive stacks
Domain
Big Data, Cloud Data Platforms (GCP), Data Engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Spark architecture and optimization, ETL processes, Hadoop ecosystem (HDFS, Hive), GCP Dataproc, Python, Java, Scala, CI/CD, Docker, Troubleshooting data systems
Preferred skills
Spark Streaming with Kafka, GCP Cloud Function/Cloud Run/Pub Sub/BigQuery
Technologies
Spark, Hadoop, Hive, Kafka, GCP (Dataproc, Cloud Function, Cloud Run, Pub Sub, BigQuery), Docker
Responsibilities
Design and implement data solutions using Cloudera technologies; Optimize data pipelines and processing workflows; Collaborate with analysts and scientists on data quality; Troubleshoot data processing and storage systems; Participate in code reviews
Seniority
Senior, hands-on IC