Software Development Engineer II - Python, PySpark, SQL
Core
Building highly-scalable, fault-tolerant distributed components and enterprise-scale applications for Fortune 500 clients.
Role type
Senior Software Development Engineer (Big Data & Distributed Systems)
Builds
Distributed computing components, APIs, data pipelines, and productionized big data jobs
Domain
Big Data, Distributed Systems, Enterprise Application Development
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Python, PySpark, Hadoop, MapReduce, HBase, ElasticSearch, SQL, Database Modeling, Data Warehousing, Unit/Integration Testing
Preferred skills
Agile methodology, complex data set handling, SQL variations
Technologies
Java, Python, PySpark, Hadoop, MapReduce, HBase, ElasticSearch, SQL
Responsibilities
Develop and evolve scalable distributed components; Design and implement APIs and integration patterns; Automate jobs and productionize big data workflows; Interface with customers to understand strategic requirements; Collaborate with engineers, data scientists, and product managers
Seniority
Mid-Senior, hands-on IC