SRE高级工程师/架构师-基础架构
Core
Ensure reliability and stability of core systems including big data, compute platforms, cloud-native, and distributed storage while driving cost optimization.
Role type
Senior SRE / Architect (Infrastructure)
Builds
Automated operations platforms for large-scale clusters and full-link monitoring systems.
Domain
Cloud-native infrastructure, big data, distributed storage
Required skills
Linux OS internals, storage architecture, network I/O models, Go/Python/Java/Shell, capacity planning, service governance, performance tuning, fault diagnosis
Preferred skills
None stated
Technologies
Linux, Go, Python, Java, Shell
Responsibilities
Design and implement automated operations solutions for large-scale clusters; Build and enhance full-link monitoring systems for observability; Collaborate with development teams to embed reliability engineering practices in architecture and release processes; Analyze performance bottlenecks and optimize critical business links; Upgrade system architectures for high availability.
Seniority
Senior, hands-on IC