Principal Data Scientist & Engineer
Core
Building the data infrastructure, models, and analytics engine for GIA, an AI-powered global HR agent that provides compliance guidance and document generation for legal and HR teams.
Role type
Principal Data Scientist & Engineer (LLM & Data Infrastructure)
Builds
LLM-based systems, data pipelines, warehousing, and product analytics for the GIA AI agent
Domain
SaaS / Global Employment Platform / Legal & HR Compliance
Deliverable
production ML models | product features | dashboards & analysis
Required skills
SQL, Python, Databricks, LLM fine-tuning, prompt engineering, RAG, traditional ML (classification, clustering, NLP), pipeline engineering, product analytics
Preferred skills
Early-stage startup experience, building product analytics from scratch, Legal/HR domain knowledge, LLM evaluation and observability, dbt, Airflow/Dagster, Spark
Technologies
Databricks, Snowflake, BigQuery, dbt, Airflow, Dagster, Spark
Responsibilities
Build and own scalable data infrastructure (pipelines, warehousing, ETL/ELT); analyze product usage and user behavior to drive decisions; build and evaluate LLM and traditional ML models; define and track key product metrics; run experiments (A/B tests, causal analysis) to measure impact; translate data insights into product strategy
Seniority
Principal, hands-on IC with strategic ownership