基础平台研发工程师(Etcd&Zk)-云存储
Core
Develop and operate ultra-large-scale Etcd and ZooKeeper clusters for cloud storage, focusing on deployment, upgrades, scaling, and architectural evolution to ensure extreme stability and high availability.
Role type
Senior Infrastructure Engineer (Distributed Systems)
Builds
Ultra-large-scale Etcd and ZooKeeper clusters for cloud storage services
Domain
Cloud Infrastructure / Distributed Systems
Required skills
Go, Java, Paxos, Raft, Linux kernel, TCP/IP, RPC, service discovery, source code debugging, capacity planning, fault recovery, observability, performance tuning
Preferred skills
Open source community contributions, AI-driven stability tools, multi-tenant isolation optimization
Technologies
Etcd, ZooKeeper, Go, Java, Linux, TCP/IP, RPC
Responsibilities
Build automated change management, capacity assessment, self-healing, and data backup systems for cluster stability; Optimize consensus protocols, throughput, latency, and handle traffic spikes; Construct observability platforms for monitoring QPS, latency, Watch/Session status, and leader election; Implement multi-active architecture, rapid recovery, and root cause analysis capabilities.
Seniority
Senior, hands-on IC
