Data Engineer (Agentic Search)
Core
Building and scaling the data platform behind agent-native search quality, ML pipelines, product analytics, and business operations for an AI search product.
Role type
Senior Data Engineer (Cloud Data Warehouse & Streaming)
Builds
Scalable data warehouse, batch and streaming pipelines, and analytics-ready datasets for internal researchers and product teams.
Domain
AI Infrastructure / Cloud Data Engineering
Deliverable
production ML models | dashboards & analysis
Required skills
Cloud data warehouse design (Snowflake/BigQuery), Medallion-style modeling, Data orchestration (Airflow), Transformation frameworks (dbt), Distributed processing (Spark/MapReduce), Streaming platforms (Kafka/Pub/Sub), Python, SQL, NoSQL databases, Production incident response, Data quality governance.
Preferred skills
Experience with multi-region production environments, Schema evolution, Cost controls, Cross-functional collaboration on ambiguous data problems.
Technologies
Snowflake, BigQuery, Airflow, dbt, Kafka, Pub/Sub, Spark, MapReduce, AWS, GCP
Responsibilities
Design and evolve scalable data models in the data warehouse; Build and maintain reliable batch and streaming pipelines; Improve observability including data quality checks, freshness monitoring, and lineage; Partner with stakeholders to deliver trustworthy datasets for product and ML analytics; Define domain objects and relationships for the search domain; Investigate and resolve production data issues; Contribute to technical standards and best practices.
Seniority
Senior, hands-on IC