Data Engineer (Python)
Core
Building and managing large enterprise data and analytics platforms, including scalable data lakes, ingestion pipelines, and hyper-scale processing clusters.
Role type
Senior Data Engineer (Python)
Builds
Scalable Smart Data Lakes, Data Ingestion Platforms, Machine Learning and NLP based Analytics Platforms, Hyper-Scale Processing Clusters, Data Mining and Search Engines
Domain
Big Data, Data Engineering, Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, PySpark, Pandas, SQL, NoSQL, Stream-processing (Kafka, Spark-Streaming), Event-driven architectures, RESTful API development
Preferred skills
Unit testing (pytest), Docker, Kubernetes
Technologies
PySpark, Pandas, Postgres, MongoDB, Elasticsearch, Kafka, Spark-Streaming, Docker, Kubernetes, pytest
Responsibilities
Create and manage end-to-end data solutions and optimal data processing pipelines for large volume, big data sets; Develop efficient pre-processing and data manipulation tasks; Implement and manage scalable data lakes and hyper-scale processing clusters; Work with data science and infrastructure teams to implement practical machine learning solutions and pipelines in production.
Seniority
Mid-Senior, hands-on IC