Site Reliability Engineer
Core
Build and maintain cloud-based infrastructure and services for OT and IoT cybersecurity platforms, ensuring availability, performance, and stability for global enterprise and government clients.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud-based cybersecurity services for operational technology (OT) and Internet of Things (IoT) infrastructures
Domain
Cybersecurity / Operational Technology (OT) / Internet of Things (IoT)
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, cloud provider management (AWS), programming (Ruby/Python/Go), Infrastructure as Code (Terraform/CloudFormation), CI/CD pipelines, AI-enabled tools integration
Preferred skills
AI-assisted tool usage, cost optimization strategies, post-incident analysis
Technologies
Kubernetes, AWS, Terraform, CloudFormation, Jenkins, GitHub Actions, Ruby, Python, Go
Responsibilities
Build and maintain cloud-based infrastructure and related services; Enhance service availability, performance, and stability; Implement automated solutions for deployment and operations; Troubleshoot issues across network, OS, and application layers; Optimize cloud service performance and costs; Conduct post-incident analysis; Perform on-call duties
Seniority
Mid-to-Senior, hands-on IC