CareerPlanSign in

大模型训练(番茄网文方向)-CQC

北京💼 Full-time🗓 2026-09-28

Core

Build and optimize large language models for content safety governance and risk assessment in the online novel (web novel) industry.

Role type

Senior IC large language model training engineer (safety alignment)

Builds

Production-ready LLMs for UGC/PGC/AIGC content moderation and compliance generation

Domain

Internet / AI Safety / Content Governance

Deliverable

production ML models

Required skills

Python (Pandas, NumPy), data engineering, prompt engineering, SFT/DPO training, Transformer/MoE architecture knowledge, distributed training (DeepSpeed, FSDP), gradient debugging

Preferred skills

Master's/PhD in CS/AI/Math, experience with Llama/Qwen/DeepSeek models, cluster training operations

Technologies

Python, Pandas, NumPy, Llama, Qwen, DeepSeek, LoRA, DPO, DeepSpeed, FSDP, Linux

Responsibilities

Convert safety standards into computable logic (regex, classification rules) and training data constraints; design and iterate prompts for content governance; execute full model training lifecycle (selection, dataset prep, fine-tuning, evaluation); resolve training errors (gradient vanishing, overfitting, OOM); collaborate with business teams to define boundary cases for safety logic.

Seniority

Senior, hands-on IC

Sourced via bytedance · Listed on CareerPlan, which tracks 844,000+ jobs from 20+ sources.