Principal Software Engineer - Spark
Core
Lead the technical roadmap and architectural vision for Cloudera Data Engineering, enabling customers to run large-scale data engineering workflows using Apache Spark, Airflow, and Iceberg across on-premises and public cloud environments.
Role type
Principal Staff Engineer (Data Infrastructure)
Builds
Cloudera Data Engineering service (cloud-native data engineering workflows)
Domain
Big Data / Data Infrastructure / Cloud-Native
Deliverable
production ML models | product features | infrastructure
Required skills
Distributed data processing systems, Cloud-native architectures, Java/Scala/C++/Python/GoLang, Containerization (Kubernetes/Docker), Apache Spark/Airflow, Public/Private Cloud (AWS/Azure/GCP/OpenShift/Rancher)
Preferred skills
Open table formats, Metadata/catalog services, Large-scale distributed systems design, Open-source contributions
Technologies
Apache Spark, Apache Airflow, Apache Iceberg, Kubernetes, Docker, AWS, Azure, GCP, OpenShift, Rancher
Responsibilities
Drive multi-year technical roadmap and architectural vision, Foster engineering excellence through mentorship and design reviews, Collaborate with cross-functional partners to deliver critical features, Work on large-scale distributed systems (hundreds to thousands of nodes)
Seniority
Principal, hands-on IC with strategy & mentorship