Sr. SRE Engineer II - EPICS, NG-SIEM (Hybrid, Sydney)
Core
Build observability, automation, and scaling systems for CrowdStrike's NG-SIEM platform to ensure reliability and performance for security event ingestion and analysis.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Observability pipelines, automated scaling solutions, and incident response tooling for the NG-SIEM platform
Domain
Cybersecurity / Large-scale distributed systems / SIEM
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Systems programming (Go, Java, Rust, C++), Scripting (Python, Bash), End-to-end observability, Incident response engineering, Capacity planning, Streaming platforms (Kafka), Infrastructure-as-code, Cross-team collaboration
Preferred skills
Hyperscaler experience, Automated remediation, Cost modeling, Cloud-native architectures, Disaster recovery planning
Technologies
Kafka, Python, Bash, Go, Java, Rust, C++, Infrastructure-as-code, CI/CD
Responsibilities
Design and maintain monitoring and synthetic test suites for the NG-SIEM pipeline; Engineer orchestrated scaling solutions to eliminate bottlenecks; Serve as subject matter expert during platform-wide incidents; Build models for end-to-end capacity forecasting and cost tracking; Transform manual SOPs into automated remediation workflows; Partner with cross-functional teams to triage SLO breaches and drive problem management
Seniority
Senior, hands-on IC