Engenheiro(a) de Dados Sênior
Core
Designing, developing, and maintaining scalable data pipelines and integrating Machine Learning models into production environments.
Role type
Senior Machine Learning Data Engineer
Builds
Scalable data pipelines, ML models (batch and near real-time), and production integrations (APIs, automated jobs)
Domain
Cloud data engineering and Machine Learning
Deliverable
production ML models
Required skills
Python, SQL, data pipeline development, feature engineering, ML libraries (Scikit-learn, TensorFlow, Pytorch), Git, cloud platform expertise (GCP or Azure), large-scale data processing (batch/streaming)
Preferred skills
Apache Spark (PySpark), MLOps practices, data security and governance, industrial/manufacturing domain experience, Agile methodologies (Scrum/Kanban)
Technologies
Google Cloud Platform (BigQuery, Dataflow, Cloud Storage, Vertex AI, Pub/Sub, Composer), Microsoft Azure, Scikit-learn, TensorFlow, Pytorch, Apache Spark, Git
Responsibilities
Design and maintain scalable data pipelines for analytics and ML; prepare and transform datasets for model training and inference; collaborate with Data Scientists to implement and support ML models; integrate ML models into production environments; ensure performance and cost efficiency in cloud environments; participate in code reviews and define technical standards; provide technical support and resolve incidents in production environments
Seniority
Senior, hands-on IC