Forward-Deployed Cheminformatician
Core
Transforming complex and heterogeneous pharmaceutical data into scalable, high-quality datasets that power next-generation scientific models and research applications.
Role type
Senior cheminformatician (drug discovery data engineering)
Builds
Modular tooling, validation frameworks, reusable pipelines, and standardized binding-data preparation protocols
Domain
Life sciences / Pharmaceutical R&D / Cheminformatics
Deliverable
production ML models
Required skills
Python, RDKit, molecular standardization, stereochemistry handling, tautomer handling, ionization management, quantitative binding assay data interpretation (KD, Ki, IC50, pIC50), data curation, pipeline development
Preferred skills
Public scientific databases, federated data environments, AI-assisted workflows, computational drug discovery platforms
Technologies
Python, RDKit
Responsibilities
Define and maintain standardized binding-data preparation protocols; Build and improve modular tooling and validation frameworks; Collaborate with researchers to interpret complex assay data and validate quality; Maintain and optimize small-molecule processing workflows; Curate and harmonize large-scale public binding-data resources; Partner with engineering and ML teams to ensure scalable data pipelines
Seniority
Mid-Senior, hands-on IC