Gen AI Engineer
Core
Analyzing and modeling organizational data to draw business insights for AI functions and decision-making.
Role type
Senior Gen AI Engineer (LLM & NLP)
Builds
Custom ML, Gen AI, NLP, and LLM models for batch and stream processing pipelines including RAG architecture.
Domain
Healthcare / Generative AI & Large Language Models
Deliverable
production ML models
Required skills
LLM development and fine-tuning, NLP, data extraction/transformation/loading, Python, SQL, Hugging Face, TensorFlow, Keras, Pytorch, Spark, GCP/AWS cloud platforms, prompt engineering, model optimization, MLOps collaboration, containerization, canary deployments
Preferred skills
vLLM, Text Generation Inference, FastAPI, model quantization (GPTQ, AWQ, bitsandbytes), GPU memory optimization, RAG architecture design, advanced cloud infrastructure (AWS EKS/ECS, GCP GKE, Azure AKS), Agile project management
Technologies
Python, SQL, Hugging Face, TensorFlow, Keras, PyTorch, Spark, vLLM, Text Generation Inference, FastAPI, GPTQ, AWQ, bitsandbytes, AWS, GCP, Azure
Responsibilities
Apply data extraction, transformation, and loading techniques to connect large data sets; Develop roadmap and strategy for NLP, LLM, and Gen AI model development; Design and develop custom ML/Gen AI/NLP/LLM models for batch and stream processing pipelines; Work with MLOps teams to create evaluation solutions for model performance; Identify and implement model optimizations; Collaborate with product teams and stakeholders for model deployment; Ensure adherence to standards and governance in ML model development
Seniority
Senior, hands-on IC