CareerPlanGet AI match score →
💼 Full-time🗓 2026-06-25

Core

Design, deploy, and maintain cloud infrastructure and containerized microservices, ensuring high availability and scalability for production services.

Role type

Mid-Level DevOps Engineer

Builds

Cloud-native infrastructure, CI/CD pipelines, and containerized microservices

Domain

Cloud infrastructure and DevOps

Deliverable

infrastructure

Required skills

Kubernetes, Terraform, CI/CD pipelines, Bash/Python/Go scripting, cloud networking, monitoring and logging, security best practices, database administration

Preferred skills

GPU-accelerated workloads, Kubernetes operators, AI/ML infrastructure, service mesh technologies, FinOps

Technologies

Kubernetes, Terraform, Helm, Jenkins, Git, Docker, Prometheus, Grafana, SQL, NoSQL

Rewrite
## About the Role We're seeking a Mid-Level DevOps Engineer to help build, scale, and maintain our cloud infrastructure. You'll work with modern DevOps tools and practices to automate deployments, optimize cloud resources, and ensure high availability of our production services across multiple environments. ## Key Responsibilities - Design, deploy, and maintain Kubernetes clusters and cloud infrastructure across multiple environments using infrastructure-as-code - Build and maintain automated CI/CD pipelines and develop automation solutions to streamline deployment processes - Deploy and manage containerized microservices in production, including specialized workloads (GPU-accelerated services) - Implement comprehensive monitoring, logging, and alerting systems with dashboards for service health and performance metrics - Participate in incident response, conduct post-incident reviews, and drive continuous reliability improvements - Manage databases, data stores, backup procedures, and disaster recovery strategies in cloud-native environments - Implement and maintain security best practices, IAM policies, and automated certificate management - Configure and optimize cloud networking, VPNs, multi-region architectures, and auto-scaling policies - Collaborate with development teams on deployment strategies while ensuring infrastructure scalability and cost-efficiency ## Required Qualifications - 2-4 years of experience in DevOps, Site Reliability Engineering, or Platform Engineering with a proven track record of managing production infrastructure at scale - Strong hands-on experience with major cloud providers (GCP, AWS, or Azure) and deep understanding of cloud-native architectures - Production experience with Kubernetes, container orchestration, and managing containerized microservices - Proficiency in infrastructure-as-code tools (Terraform, CloudFormation, or similar) and GitOps practices - Experience building and maintaining automated CI/CD pipelines for continuous delivery - Strong scripting and automation skills in Bash, Python, or Go - Solid networking fundamentals including TCP/IP, DNS, load balancing, VPNs, and multi-region architectures - Experience with monitoring, logging, and observability platforms (Prometheus, Grafana, or similar) - Knowledge of security best practices, IAM policies, and compliance requirements - Familiarity with database administration and data infrastructure in cloud environments - Excellent troubleshooting abilities with experience in incident response and on-call rotations - Strong communication, collaboration, and documentation skills with ability to work independently in fast-paced environments ## Nice to Have - Experience with specialized workloads (GPU computing, real-time media processing) - Knowledge of real-time communication technologies (WebRTC, SIP, media servers) - Experience with Kubernetes operators and custom resource definitions - Understanding of AI/ML infrastructure and deployment patterns - Experience designing multi-region, highly available architectures - Background in platform engineering or internal developer platforms - Relevant certifications (CKA, CKAD, cloud provider certifications) - Contributions to open-source DevOps or infrastructure projects - Experience with service mesh technologies (Istio, Linkerd) - Knowledge of FinOps and cloud cost optimization strategies ## Technology Stack Our infrastructure leverages modern, industry-standard tools and platforms: - Cloud & Infrastructure: Kubernetes, Terraform, Helm, major cloud platforms - CI/CD & Development: Jenkins, Git, Docker, container registries - Observability: Prometheus, Grafana, centralized logging - Data Layer: SQL and NoSQL databases, caching systems - Networking: Ingress controllers, service mesh, certificate automation
Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Wellfound ↗