Staff Software Development Engineer - Data Platforms
Core
Design, build, and optimize critical data infrastructure handling hundreds of millions of requests per minute, processing billions of daily events, and managing petabytes of storage to support AI-powered experiences and autonomous agents.
Role type
Staff Software Development Engineer (Data Platforms)
Builds
Scalable data ingestion services, real-time event collection systems, and data foundations for AI model training and autonomous agents.
Domain
Consumer media streaming, AI/ML infrastructure, distributed systems
Deliverable
production ML models | infrastructure
Required skills
Python, Java, Scala, or Go; Apache Spark; Apache Flink; Apache Kafka; AWS (S3, EMR, Kinesis, Glue, Redshift); Snowflake, BigQuery, Redshift, Delta Lake, Iceberg, Hudi; Apache Airflow or Prefect; PostgreSQL, MySQL, Cassandra, DynamoDB; dimensional and data vault modeling
Preferred skills
Experience in consumer-facing or streaming industry; building data pipelines for ML training and inference; feature stores; data governance
Technologies
Kafka, S3, HDFS, Spark, Flink, Airflow, Prefect, AWS, GCP, Azure, Snowflake, BigQuery, Redshift, Delta Lake, Iceberg, Hudi, PostgreSQL, MySQL, Cassandra, DynamoDB
Responsibilities
Architect and design new services and migrations; build and maintain real-time event collection systems; create data pipelines for AI model training and autonomous agents; detect and fix data quality issues; improve existing systems and optimize unit economics; own the full development lifecycle from ideation to deployment; mentor junior engineers
Seniority
Staff, technical leadership & strategic architecture