Senior Site Reliability Engineer
Core
Design, deploy, maintain, and optimize tool sets for Information Security teams to defend against adversaries and ensure production systems operate smoothly within uptime objectives.
Role type
Senior Site Reliability Engineer (Cyber Security Engineering)
Builds
Hybrid cloud production containerization service offering, CI/CD pipelines, and distributed systems for security operations.
Domain
Cyber Security / Information Security / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Cloud Platform administration (AWS, GCP, Azure), Observability implementation, CI/CD pipeline management, Microservices and containerized application deployment, Distributed system architecture, Linux and Windows administration, GitOps practices, Infrastructure as Code (Terraform, ArgoCD, Helm), Scripting (Python, Go, Elixir), Database management (NoSQL, Relational)
Preferred skills
SPIRE/SPIFFE implementation, Serverless architecture design, Data management and pipeline technologies (Kafka, Spark, Flink), OpenTelemetry/Prometheus/Grafana proficiency, AWS Cloud Practitioner or Azure AZ-900 certification
Technologies
AWS, GCP, Azure, Kubernetes, Amazon ECS, Docker, Terraform, ArgoCD, Helm, GitHub Actions, Bamboo, Jenkins, Azure DevOps, OpenTelemetry, Prometheus, Grafana, Apache Storm, Kafka, Flink, Spark, Hadoop, Python, Go, Elixir, Git, Mercurial, Subversion
Responsibilities
Architect new and existing systems to enhance performance, reliability, and scalability; Build and iterate over CI/CD pipelines; Manage, develop, design, and deploy microservice and containerized applications; Implement strong security controls in distributed systems; Coordinate with engineers to automate deployments and configurations; Abstract complexity of Observability implementation by writing scalable automation; Identify opportunities for improvement around observability and process; Standardize and develop alerts/notifications and response to monitoring tools; Work alongside application teams to implement Observability in day-to-day operations; Contribute to post-mortems and provide root cause analysis; Design and implement standards, policies, and procedures for automation and integrations.
Seniority
Senior, hands-on IC