Data Engineer- AWS
Core
Design and build cloud-native data pipelines, orchestrate automation workflows, and ensure data integrity for real-time insights and enterprise operations.
Role type
Senior IC data engineer (AWS cloud-native)
Builds
Scalable, automated, and test-driven data pipelines and automation workflows
Domain
Cloud data engineering (AWS) + Big Data technologies
Deliverable
production ML models | product features | infrastructure
Required skills
Python, SQL, Scala, Java, PySpark, Hadoop, Hive, EMR, Kafka, Spark, NoSQL (MongoDB, Cassandra), Data warehousing (Redshift), UNIX/Linux shell scripting, SQL performance tuning, process orchestration (Airflow, AWS Step Functions, Luigi, KubeFlow)
Preferred skills
Machine Learning workflows, Data-as-a-Service platforms, API design and deployment
Technologies
AWS, Redshift, Kafka, Spark, Hadoop, Hive, EMR, MongoDB, Cassandra, Airflow, AWS Step Functions, Luigi, KubeFlow
Responsibilities
Develop, test, deploy, orchestrate, monitor, and troubleshoot cloud-based data pipelines; Collaborate with data scientists, architects, and stakeholders to integrate data from internal and external sources; Research and experiment with batch and streaming data technologies; Contribute to the definition and continuous improvement of data engineering processes; Ensure data integrity, accuracy, and security across corporate data assets; Apply software development practices including Git-based version control, CI/CD, and release management
Seniority
Mid-Senior, hands-on IC