Sr Machine Learning Engineer
Core
Design and develop scalable, secure, and reliable data pipelines and ingestion solutions that power knowledge layers and assistant experiences via generative AI for Manufacturing Applications.
Role type
Senior Data Engineer (Generative AI & Manufacturing)
Builds
Scalable data pipelines, data integration frameworks, metadata-driven architectures, and AI-driven insights for Manufacturing and Operations.
Domain
Biotechnology/Pharma Manufacturing
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Databricks, PySpark, Scala, SQL, AWS, workflow orchestration, big data processing, streaming technologies (Apache Kafka, Debezium), data modeling, LLMs, vector stores, RAG, knowledge graphs
Preferred skills
AI-assisted code development tools, API development, SQL/NoSQL databases, OLAP/OLTP performance tuning, manufacturing data sources (SCADA, Data Historian)
Technologies
Databricks, PySpark, Scala, SQL, AWS, Apache Spark, Apache Kafka, Debezium, PostgreSQL, MySQL, SQL Server, MongoDB, Git, Jenkins, Maven, JIRA, Confluence
Responsibilities
Design and maintain complex ETL/ELT data pipelines; Champion ML/LLM feature engineering including embeddings and vector DBs; Migrate and deploy complex data across systems; Implement secure access, logging, privacy controls, and governance; Ingest and transform structured and unstructured data from various sources; Ensure data integrity through quality checks and monitoring; Innovate and implement new tools for efficient data processing; Automate tasks and develop reusable frameworks.
Seniority
Senior, hands-on IC