Remote Senior Platform Engineer
Core
Design, deploy, and support secure, scalable Azure cloud infrastructure and automation in production.
Role type
Senior IC Platform Engineer (Cloud Infrastructure)
Builds
Production cloud infrastructure, CI/CD pipelines, containerized workloads, and automation scripts.
Domain
Cloud Infrastructure / DevOps
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Azure cloud infrastructure, Terraform, Infrastructure as Code, Python, PowerShell, Bash, Kubernetes, CI/CD pipelines, cloud networking, cloud security, cloud governance, cloud monitoring, observability, root-cause analysis, disaster recovery, business continuity, cost optimization, high availability, autoscaling, automated remediation, architecture documentation, change management, incident response, capacity planning, service health monitoring, technical support materials creation, production readiness, knowledge transfer, on-call support.
Preferred skills
AI
Technologies
Azure, Bash, CI/CD, Cloud, DevOps, Support, Kubernetes, PowerShell, Python, Security, Terraform, AI, LESS
Responsibilities
Design, deploy, and support secure, scalable cloud infrastructure in production; Automate infrastructure provisioning and configuration using Infrastructure as Code practices; Build and maintain cloud networking, compute, storage, database, container, and integration services; Develop and manage CI/CD pipelines for automated infrastructure and application deployments; Administer containerized workloads and Kubernetes environments; Support cloud operations through troubleshooting, incident response, root-cause analysis, and change management; Apply cloud security, governance, identity, compliance, and cost-management controls; Design highly available, resilient environments that support disaster recovery and business continuity; Implement monitoring, alerting, autoscaling, and automated remediation; Create scripts and automation to reduce manual work and improve operational efficiency; Monitor cloud performance, availability, capacity, and service health; Prepare and maintain architecture documentation, operational procedures, and technical support materials; Collaborate with engineering and operations teams on deployments, knowledge transfer, and production readiness; Participate in on-call support for critical production environments as needed.