Data Engineer I
Core
Design, build, and maintain scalable data infrastructure and pipelines to support clinical, operational, and strategic decision-making in healthcare.
Role type
Data Engineer I
Builds
Scalable data lakes, warehouses, and ETL/ELT pipelines for healthcare data.
Domain
Healthcare / Big Data Engineering
Deliverable
production ML models | infrastructure
Required skills
SQL, Python, AWS (EC2, EMR, RDS, Redshift, Glue, DynamoDB), Hadoop, Spark, Kafka, NoSQL databases, data pipeline tools, schema design, indexing, partitioning, CI/CD, data validation, anomaly detection, HL7, FHIR
Preferred skills
Advanced SQL, Python optimization, metadata management, workload orchestration, stream processing, project management
Technologies
AWS, Snowflake, Redshift, BigQuery, Hadoop, Spark, Kafka, SQL, NoSQL, Python, Java, C++, Scala
Responsibilities
Design and maintain optimal data pipeline architecture for structured and unstructured healthcare data; Build scalable ETL/ELT pipelines using SQL and AWS big data technologies; Create and maintain data lakes, warehouses, and marts; Develop reusable components for reporting and dashboarding tools; Implement secure and compliant data architectures following HIPAA; Collaborate with stakeholders to define KPIs and metrics; Provide mentorship to junior data engineers.
Seniority
Junior to Mid-level, hands-on IC