Senior Cloud Performance Engineer
Core
Building and operating scalable, fault-tolerant, distributed systems for ClickHouse Cloud, focusing on performance limits, benchmarking, and chaos engineering.
Role type
Senior distributed systems performance engineer
Builds
Elastic, limitless scale, high-performance, serverless ClickHouse Cloud platform
Domain
Cloud infrastructure + distributed databases (OLAP)
Deliverable
production ML models | infrastructure
Required skills
distributed systems architecture, database benchmarking, test automation, system engineering, performance analysis, capacity management, Go/C/C++/Java, concurrency, multithreading, Kubernetes, public cloud infrastructure (AWS/GCP/Azure)
Preferred skills
leading large scope technical projects, production debugging, chaos engineering techniques
Technologies
ClickHouse, Kubernetes, AWS, GCP, Azure, Go, C/C++, Java
Responsibilities
Benchmark system and database performance; troubleshoot and debug applications and server errors; recommend configuration tuning for performance bottlenecks; partner with core development, cloud, and security teams to improve ClickHouse Cloud performance; plan and drive Chaos Engineering initiatives; develop and manage tools for chaos experiments; observe running systems to prioritize innovative disruption methods
Seniority
Senior, hands-on IC