ETL Developer + Snowflake + Hadoop + Informatica
Core
Design, develop, and maintain scalable ETL pipelines for large-scale data processing, integration, and enterprise analytics using Hadoop, Spark, and cloud platforms.
Role type
Senior ETL Developer (Data Engineering)
Builds
End-to-end batch and real-time data pipelines, data integration solutions, and optimized data workflows.
Domain
Enterprise Data Engineering, Cloud Data Platforms, Big Data
Required skills
ETL tools (Informatica, Talend), Hadoop ecosystem (HDFS, Hive, Pig, Sqoop), Apache Spark (PySpark, Spark SQL), Python/Java/Scala, SQL, Relational databases (Oracle, PostgreSQL, MySQL)
Preferred skills
Cloud data migration tools, NoSQL databases (DynamoDB, Cassandra), Data governance and metadata management, AI/ML frameworks (LangChain, Hugging Face)
Technologies
Spark, Hive, Pig, AWS (EMR, S3, Glue, Redshift), Azure, GCP, Jenkins, GitHub Actions, Terraform, CloudFormation, Ansible
Responsibilities
Design and optimize ETL pipelines for large-scale data processing; Build batch and real-time workflows using Spark and Hadoop; Convert business requirements into high-performance data solutions; Perform performance tuning and debugging of data workflows; Ensure data quality, security, and compliance; Automate data ingestion and deployment pipelines; Monitor workflows and troubleshoot incidents; Implement data governance and metadata management standards.
Seniority
Senior, hands-on IC