Data Engineer
Core
Build and maintain data pipelines and infrastructure to support Data Science projects, enabling storage, streaming, and processing of large-scale data for client challenges.
Role type
Data Engineer supporting Data Science initiatives
Builds
Data pipelines, ML infrastructure, and scalable data solutions for clients
Domain
Data Engineering, Big Data, Machine Learning Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, ETL development, relational databases (PostgreSQL, SQLServer), distributed computing (Hadoop, Spark)
Preferred skills
Additional programming languages, NoSQL paradigms (Wide-column, Key-value), Cloud platforms (GCP, AWS, Azure), Machine Learning frameworks (TensorFlow, Torch)
Technologies
Python, Hadoop, Spark, TensorFlow, Torch, PostgreSQL, SQLServer, GCP, AWS, Azure
Responsibilities
Develop ETL scripts in Big Data ecosystems, implement ML infrastructure for model training and deployment, contribute to architectural and governance choices for data scaling, build server-side data processing tools and APIs
Seniority
Mid-level, hands-on IC