Data Engineer
Core
Build and maintain data pipelines and infrastructure to support Data Science projects, enabling storage, streaming, and processing of large-scale data for clients.
Role type
Data Engineer supporting Data Science (via careerplan.io/jobs/744000153149349-data-engineer-at-sia)
Builds
Data pipelines, ML infrastructure, and scalable data services for client projects
Domain
Data Engineering, Big Data, Cloud Infrastructure
Deliverable
production ML models
Required skills
Python, SQL, PostgreSQL, distributed computing (Hadoop/Spark), data processing, API development
Preferred skills
Additional programming languages, NoSQL databases (Wide-column, Key-value), Cloud platforms (GCP/AWS/Azure), Machine Learning frameworks (TensorFlow/Torch)
Technologies
Python, PostgreSQL, SQLServer, Hadoop, Spark, TensorFlow, Torch, GCP, AWS, Azure
Responsibilities
Develop ETL scripts in Big Data ecosystems, implement ML infrastructure for model training and deployment, architect cloud storage and compute solutions, contribute to data governance and technology choices
Seniority
Mid-level, hands-on IC
