Site Reliability Engineer
Core
Maintain uptime, scalability, and security for AWS-hosted MERN applications and backend data architecture in a HIPAA-regulated healthcare environment.
Role type
Senior Site Reliability Engineer (Cloud & Data)
Builds
Serverless workloads, distributed data orchestration pipelines, and immutable infrastructure.
Domain
Healthcare technology, Cloud Infrastructure, Data Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
AWS core services (Lambda, ECS/EKS, EMR/Glue, EC2, VPC, IAM, CloudWatch), Python (PySpark), Node.js, distributed data orchestration, ETL tooling, message queues (SQS/SNS, RabbitMQ), MySQL administration, Terraform/OpenTofu, incident management, HIPAA compliance.
Preferred skills
Securing sensitive data at rest and in transit, managing memory leaks and cold starts in serverless environments.
Technologies
AWS Lambda, ECS, EKS, EMR, Glue, EC2, VPC, IAM, CloudWatch, PySpark, Node.js, SQS, SNS, RabbitMQ, MySQL, Athena, Terraform, OpenTofu, MERN, REST
Responsibilities
Manage and tune serverless workloads on AWS Lambda; monitor and troubleshoot scheduled PySpark workflows; participate in on-call rotation for application outages and data bottlenecks; ensure HIPAA, SOC2, and HITRUST compliance; build automation for infrastructure provisioning and data recovery; develop dashboards for monitoring Node.js and PySpark performance; lead blameless post-mortems.
Seniority
Senior, hands-on IC
