Member of Technical Staff (Software Engineer, Data Platform)
Core
Designing and operating large-scale batch and streaming data pipelines, orchestration systems, and self-serve platforms to power product features, AI workloads, and analytics at Perplexity.
Role type
Senior/Staff Software Engineer (Data Platform)
Builds
End-to-end data lifecycle systems including ingestion, processing, storage, and serving layers for AI and product teams.
Domain
Data Engineering / AI Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Large-scale batch and streaming data processing, data orchestration (Airflow/Dagster), Python, backend languages (Go/TypeScript), systems thinking for reliability and latency, ML/AI workflow support, data quality and lineage tooling, internal platform ownership.
Preferred skills
Experience with Databricks, Snowflake, Spark, Kafka, Flink, Iceberg, Delta Lake, ClickHouse.
Technologies
Databricks, Snowflake, Spark, Kafka, Kinesis, PubSub, Airflow, Dagster, dbt, Iceberg, Delta Lake, ClickHouse
Responsibilities
Design and operate large-scale batch and streaming data pipelines; Build event-driven and streaming systems for real-time ingestion and transformation; Lead the architecture of data orchestration and observability; Set guarantees for data correctness, freshness, and recoverability; Build self-serve data platforms for engineers and data scientists; Improve developer experience through abstractions and standards; Drive architectural decisions across storage, compute, and APIs; Mentor engineers and review designs.
Seniority
Senior/Staff, hands-on IC with architectural leadership