CareerPlanSign in

Member of Technical Staff, ML Engineer

💼 Full-time🗓 2026-10-02

Core

Building the inference APIs, batch/compute systems, and services that make the Omnii genome language model fast, reliable, and cheap at genome scale for scientists and partners.

Role type

Senior IC ML Engineer (Inference & Distributed Systems)

Builds

Inference APIs, SDKs, documentation, and compute/orchestration layers for the Omnii model.

Domain

Generative AI / Genomics / Distributed Systems

Deliverable

production ML models

Required skills

Python, typed languages, containers, infrastructure as code, API design, distributed systems, batch job systems, observability, cost/latency optimization

Preferred skills

GPU inference internals (vLLM, TensorRT-LLM, Triton), enterprise VPC/on-prem/air-gapped deployment, genomics/computational biology

Technologies

Python, vLLM, TensorRT-LLM, Triton

Responsibilities

Build real-time and batch inference APIs for genome-scale workloads; design compute orchestration with autoscaling and fault tolerance; develop SDKs and documentation for external teams; optimize inference cost and latency; package platform for on-prem and air-gapped environments; collaborate with scientists to shape API workloads.

Seniority

Senior, hands-on IC

Sourced via wellfound · Listed on CareerPlan, which tracks 928,000+ jobs from 20+ sources.