Data Engineer
Core
Design and build data pipelines from scratch to handle petabytes of logs, events, and model traces, creating a clean environment for production, testing, and research workloads.
Role type
Data Engineer
Builds
Internal data tooling, structured datasets, data pipelines, and analytics dashboards
Domain
AI Safety, Large Language Models (LLMs), Cloud Data Warehousing
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Python, SQL, Web Scraping, PostgreSQL, Cloud Data Warehousing (ClickHouse, BigQuery, Redshift, Snowflake)
Preferred skills
Metabase, Dataset Versioning, dbt, ETL tools (Fivetran, Airbyte, dltHub), AWS (Athena, Glue, S3), AI-assisted coding
Technologies
Python, SQL, PostgreSQL, ClickHouse, BigQuery, Redshift, Snowflake, Metabase, dbt, Fivetran, Airbyte, dltHub, AWS, Athena, Glue, S3
Responsibilities
Own and evolve internal data tooling; Scrape external sources and turn raw data into structured datasets; Build and maintain pipelines to transform and version data; Create new data products addressing business needs; Diagnose and fix pipeline issues; Ship analytics and dashboards
Seniority
Mid-Senior, hands-on IC