Software Development Engineer, Open Data Analytics - Engines
Core
Design, implement, and optimize core components of query engines (Apache Spark, Trino) and open table formats (Apache Iceberg, Hudi, Delta) for serverless cloud environments.
Role type
Senior IC backend engineer (distributed systems & query engines)
Builds
High-performance, scalable data lake workloads and analytics services for AWS customers
Domain
Cloud computing, Big Data, Distributed Systems
Deliverable
production ML models | product features
Required skills
Distributed systems architecture, Query engine optimization, Open source community collaboration, System design & scaling, Performance tuning, Security hardening, Automation & testing, Mentorship
Preferred skills
Full SDLC experience, Compiler development, Algorithm design
Technologies
Apache Spark, Trino, Apache Iceberg, Hudi, Delta, Java, C++
Responsibilities
Develop and optimize core components of query engines and open table formats; Design and implement innovative solutions to improve feature capabilities and stability; Collaborate with the open-source community to drive improvements; Ensure data consistency and durability for large-scale data lake workloads; Mentor and train other team members on design techniques and coding best practices
Seniority
Senior, hands-on IC with leadership responsibilities