Staff Software Engineer, Data Quality and Governance
Core
Lead a team building large-scale data discovery, metadata, and catalog platforms to ensure data trustworthiness and findability across Stripe.
Role type
Staff Software Engineer, Data Quality and Governance
Builds
Data Catalog, Knowledge Graph Service, dataset tiering, governance standards, lineage tracking, and quality scoring systems
Domain
Financial infrastructure / Data Engineering / Distributed Systems
Deliverable
production ML models | product features | infrastructure
Required skills
Distributed systems fundamentals, Backend development (Scala/Java/Go), SQL, Team leadership, Data pipeline management, Full-stack web application development, Data quality analysis, AI/LLM and Agents integration, Project lifecycle management
Preferred skills
Iceberg, Kafka, Change Data Capture, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, AWS Cloud, Open-source contributions, Data mart/warehouse creation, Cross-functional product collaboration
Technologies
Iceberg, Kafka, Flink, Spark, Airflow, Hive Metastore, Pinot, Trino, AWS Cloud
Responsibilities
Lead technical outcomes and mentor engineers; Build and operate large-scale data discovery and catalog platforms; Develop subject matter expertise and manage SLAs for data pipelines and web apps; Collaborate to create canonical datasets and data warehouses; Leverage AI/LLMs to analyze high-quality data on ambiguous problems; Drive execution of key data initiatives from planning to delivery
Seniority
Staff, hands-on IC with team leadership
