Sr Technical Consultant - Azure Cloud, DevOps, Sql & Site Reliability Engineering(SRE)
Core
Senior Site Reliability Engineer responsible for troubleshooting, incident management, and automation of the BYX Azure cloud platform supporting enterprise applications.
Role type
Senior IC Site Reliability Engineer (Azure/Kubernetes)
Builds
Shared cloud provisioning, compute, storage, observability, and runtime capabilities for enterprise applications.
Domain
Cloud Infrastructure / DevOps / Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Azure cloud services, Kubernetes troubleshooting, container registry management, observability (logs/metrics/traces), Elastic Stack, Linux administration, scripting (Bash/Python/PowerShell), CI/CD workflows, networking (DNS/TLS/CIDR), incident response, root cause analysis, capacity planning.
Preferred skills
MongoDB/Atlas, Redis, SQL, NFS, Infrastructure as Code (Terraform/Bicep/ARM), Helm, disaster recovery workflows.
Responsibilities
Review and act on incidents and infrastructure requests; own L2/L3 cloud-platform issues from triage to closure; troubleshoot Kubernetes pods, deployments, and resource constraints; diagnose container image-pull and registry-authentication failures; investigate autoscaling issues; support Azure resource provisioning and lifecycle workflows; diagnose networking and security connectivity problems; trace logs and telemetry through observability pipelines; support data and storage service operations; manage certificates and credentials; support regional releases and disaster-recovery workflows; develop automation to improve platform reliability; maintain runbooks and diagnostic procedures; participate in capacity planning and on-call rotation.
Seniority
Senior, hands-on IC