Cloud Data Engineer
Core
Design, build, and evolve modern cloud data architectures, including scalable ETL/ELT pipelines and data lake/warehouse management, supporting Data Science and AI initiatives.
Role type
Cloud Data Engineer
Builds
Scalable data pipelines, Data Warehouse, and Data Lake ecosystems on GCP and AWS.
Domain
Cloud Data Engineering, Data & AI
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Cloud data architecture, ETL/ELT pipeline development, Data Warehouse management, Data Lake management, SQL, Python, Go, Node.js, C/C++, Workflow orchestration, CI/CD, Infrastructure as Code
Preferred skills
Generative AI concepts, LLM orchestration, Vector Database, Data security and encryption, Terraform, Cloud certifications
Technologies
Google Cloud Platform (GCP), AWS, BigQuery, Cloud Storage, Cloud Dataflow, Cloud Pub/Sub, Cloud Composer, Amazon S3, AWS Glue, Amazon Athena, Amazon Redshift, AWS Lambda, Apache Airflow, LangChain, LlamaIndex, CrewAI, AutoGen, Terraform
Responsibilities
Design and develop scalable ETL/ELT pipelines on GCP and AWS; Implement batch and real-time data acquisition and processing; Manage and optimize Data Warehouse and Data Lake ecosystems; Collaborate with Data Science and AI teams to integrate ML models and support GenAI/Agentic AI solutions; Ensure data quality, security, and compliance throughout the data lifecycle; Automate deployment and operational processes using Infrastructure as Code and CI/CD pipelines; Participate in the evolution of cloud data architectures for clients.
Seniority
Mid-level (2-3 years experience)