Member of Technical Staff, Evals & Post-Training Product
Core
Building products and workflows that connect evaluation and post-training into a continuous loop for developers to improve generative AI models.
Role type
Product Engineer specializing in LLM evaluations and post-training
Builds
Internal evaluation tooling, fine-tuning product experiences (SFT, RFT), and the Eval Protocol SDK
Domain
Generative AI infrastructure, LLM inference, and model optimization
Deliverable
product features
Required skills
LLM evaluation design, post-training methods (SFT, RFT), full-stack product engineering, GenAI lifecycle understanding, user-centric product development
Preferred skills
Domain-specific evaluation experience (medical, legal, coding), open-source contributions, inference/hardware optimization knowledge, startup environment experience
Technologies
APIs, SDKs, backend systems, web app, Eval Protocol SDK
Responsibilities
Design and scale internal evaluation tooling for model quality measurement; Build and improve user-facing workflows for fine-tuning custom models; Partner with customers and stakeholders to convert bespoke workflows into productized solutions
Seniority
Mid-level to Senior Product Engineer