Principal Software Engineer - Big Data & Java
Core
Design, develop, and maintain large-scale data platforms and pipelines using Java microservices to support healthcare data analytics and AI initiatives.
Role type
Principal Software Data Engineer (Big Data & Java)
Builds
Scalable distributed systems, batch and real-time data pipelines, and cloud-native data solutions.
Domain
Healthcare technology, Big Data, Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Microservices, Apache Hudi, Apache Trino, Azure ADLS, Event-driven architectures, Distributed systems, Cloud platforms (AWS/Azure/GCP), Data governance, Observability, CI/CD, Lakehouse architectures
Preferred skills
Mentoring engineers, Schema management, Automated testing frameworks (dbt, Great Expectations), Performance tuning
Technologies
Java, Apache Hudi, Apache Trino, Azure ADLS, AWS, Azure, GCP, HDFS, Databricks, Spark
Responsibilities
Lead design and implementation of scalable distributed systems; Engineer and optimize data pipelines; Collaborate cross-functionally with product, analytics, and AI teams; Advance modernization of event-driven architectures; Drive adoption of data governance and observability best practices; Embed data quality checks in processing pipelines; Establish robust observability for data pipelines.
Seniority
Principal, hands-on IC with mentorship