Staff Engineer Cloud Scalability
Core
Lead optimization, monitoring, and tuning of cloud infrastructure to improve platform availability and scalability for a global virtual workspace platform.
Role type
Staff Cloud Scalability Engineer
Builds
High-availability, scalable cloud infrastructure for a secure mobility SaaS platform
Domain
Cybersecurity SaaS / Cloud Infrastructure
Deliverable
infrastructure
Required skills
Cloud systems architecture, capacity planning, autoscaling, performance optimization, observability, SLO/SLA enforcement, incident response, cost optimization, microservices, containerization, LLM tools
Preferred skills
AWS expertise, modern tooling (CloudWatch, X-Ray, OpenTelemetry, Splunk)
Technologies
AWS, CloudWatch, X-Ray, OpenTelemetry, Splunk
Responsibilities
Lead capacity planning, autoscaling, and performance optimization across the application; Define and enforce best practices for scalability, reliability, observability, and infrastructure resilience; Conduct architectural reviews and propose improvements to enhance performance and cost efficiency; Partner with engineering teams on the implementation of micro service infrastructure; Implement and improve metrics systems to diagnose performance bottlenecks; Partner with engineering teams to help enforce SLOs/SLAs; Lead incident response initiatives for high-severity cloud infrastructure issues; Provide solutions to enable production teams to measure and predict cloud costs against resource utilization; Participate in on-call, production assistance, and escalation schedules
Seniority
Staff, hands-on IC