Big Data Engineer
Core
Design, build, and maintain large-scale distributed data processing systems and pipelines.
Role type
Big Data Engineer
Builds
Large scale analytic systems and data pipelines
Domain
Big Data, Distributed Systems, Cloud Computing
Deliverable
production ML models | infrastructure
Required skills
Java, Apache Spark, Hadoop, distributed system design, data pipelining, machine learning algorithms, software design patterns, OO design principles, UML, data modeling, map/reduce computation, Linux clusters, Agile/Scrum methodology
Preferred skills
CUDA, threads, MPI, research-oriented mindset, fast learning
Technologies
Java, Spark, Hadoop, HDFS, Linux, Cloud technologies
Responsibilities
Design and implement distributed computing systems; build large scale applications using software design patterns; convert raw data into structured sets for productization; work with data warehousing and parallel processing of large datasets; apply modern development methodologies in a fast-paced technical environment
Seniority
Mid-level (2+ years experience)