大数据开发工程师-流式计算方向
Core
Design and develop real-time streaming data systems and distributed infrastructure for large-scale recommendation engines, ensuring stability, high availability, and performance optimization.
Role type
Senior IC big data engineering specialist (streaming computing)
Builds
Real-time streaming data systems, scalable storage and compute models, and distributed infrastructure for recommendation platforms.
Domain
Internet / Recommendation Systems / Big Data Streaming
Deliverable
production ML models
Required skills
Flink (DataStream, FlinkSQL, Checkpoint, State), Kafka, RocketMQ, Java, C++, Scala, Python, Data Lake technologies (Hudi, Iceberg, DeltaLake), distributed systems architecture, performance tuning, troubleshooting.
Preferred skills
PB-level data processing experience, source code reading experience for Flink/Kafka/RocketMQ, storage system experience (HBase, Cassandra, RocksDB), Spark, Kudu, YARN, K8S.
Responsibilities
Design and implement real-time streaming data systems for large-scale recommendation engines; build flexible, scalable, and high-performance storage and compute models; troubleshoot production systems and implement mechanisms/tools for stability; develop industry-leading streaming computing frameworks and distributed systems.
Seniority
Senior, hands-on IC