Staff Software Engineer - Apache Spark
Core
Architect and build next-generation, enterprise-grade distributed data processing systems running Apache Spark on thousands of nodes for the world's largest companies.
Role type
Staff Software Engineer (Distributed Systems / Big Data)
Builds
Cloudera Data Platform (CDP) control and data planes, specifically focusing on Iceberg and Spark components.
Domain
Big Data, Distributed Systems, Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Scala, distributed systems design, SQL planners and optimizers, fault tolerance, large-scale system debugging, infrastructure tooling
Preferred skills
Apache Spark development, Apache Iceberg, open-source contributions, performance optimization, scheduling
Technologies
Apache Spark, Apache Iceberg, Scala, Java, Python
Responsibilities
Architect and implement next-generation features for the Data Engineering Experience; contribute to Apache Spark open-source; develop high-performance features using modern data platforms; debug complex system-level deployment issues; improve internal infrastructure and tooling; collaborate with distributed teams to drive product vision.
Seniority
Staff, hands-on IC with strategic influence