Staff Software Engineer - Replication Manager
Core
Design, develop, and maintain enterprise-grade data replication solutions enabling seamless data movement across hybrid and multi-cloud environments for Fortune 500 companies.
Role type
Staff Software Engineer (Distributed Systems & Data Replication)
Builds
Scalable replication services, APIs, and microservices for the Replication Manager platform.
Domain
Big Data, Cloud Infrastructure, Distributed Systems
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Scala, Python, Apache Hadoop ecosystem (HDFS, Hive, HBase, YARN), Apache Iceberg, Delta Lake, Kafka, Pulsar, AWS, Azure, GCP, Docker, Kubernetes, distributed systems architecture, data consistency models, CAP theorem, microservices, API design, security protocols, data governance frameworks, agile SDLC, CI/CD, automated testing, Git, Prometheus, Grafana, ELK stack
Preferred skills
Apache Ranger, Apache Atlas, enterprise backup and disaster recovery solutions, data migration, ETL pipeline development, open-source big data contributions, customer-facing support
Technologies
HDFS, Hive, HBase, Apache Iceberg, Kafka, Pulsar, AWS S3, Azure ADLS Gen2, Docker, Kubernetes, Prometheus, Grafana, ELK stack, Apache Atlas
Responsibilities
Design and implement scalable replication services across big data technologies; lead complex data migration initiatives between on-premises and cloud; build robust APIs and microservices; design fault-tolerant petabyte-scale distributed systems; ensure data security and governance compliance; drive technical decisions for new features; guide junior engineers on best practices
Seniority
Staff, hands-on IC with mentorship