Senior Cloud / AWS Infrastructure Engineer
Core
Build, operate, secure, and evolve a large-scale AWS environment supporting a global consumer platform, with a major focus on cross-region disaster recovery and cloud security.
Role type
Senior Cloud / AWS Infrastructure Engineer
Builds
Large-scale AWS infrastructure, disaster recovery systems, and security controls for a global consumer platform
Domain
Cloud Infrastructure / AWS
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
AWS infrastructure design, Terraform (module design, state management), AWS Organizations/Control Tower, IAM (policy design, federation, least-privilege), container orchestration (ECS/Kubernetes), relational databases, Python/Bash scripting, legacy infrastructure migration, security incident response, monitoring/alerting (Prometheus, Grafana, CloudWatch), automation tooling development
Preferred skills
AWS account-vending technologies (Account Factory, Landing Zone Accelerator), IAM Identity Center/SAML/OIDC federation, cross-region disaster recovery testing, Amazon Aurora at scale, AWS security tools (Security Hub, GuardDuty, Config, CloudTrail), AWS Backup at organizational scale, Cloudflare/CDN/WAF, data/AI workloads (Redshift, MWAA, Bedrock), distributed team collaboration, AWS certifications
Technologies
AWS, Terraform, Python, Bash, ECS, Kubernetes, Amazon Aurora, Prometheus, Grafana, Alertmanager, CloudWatch, AWS Control Tower, AWS Secrets Manager, SSM Parameter Store, AWS Config, Security Hub, GuardDuty, CloudTrail, AWS Backup, Cloudflare, Redshift, MWAA, Bedrock, Claude, Claude Code, Cowork
Responsibilities
Extend and maintain versioned Terraform module libraries; Design and maintain VPC architectures and cross-account networking; Support migration of legacy infrastructure to standardized AWS structures; Design and implement cross-region disaster recovery for critical platforms; Establish recovery objectives (RPO/RTO) with business stakeholders; Standardize organization-wide backup policies; Conduct scheduled disaster recovery tests and game days; Design and strengthen IAM policies and federated SSO; Develop automation and operational tooling using Python and Bash; Maintain and evolve AWS Control Tower and landing zones; Troubleshoot production issues across infrastructure layers
Seniority
Senior, hands-on IC