About the job
Core
Generate, review, and refine high-quality Korean NLP datasets to improve foundation model inference capabilities and evaluate model performance.
Role type
Korean NLP Data Engineer / Data Annotator
Builds
Training and evaluation datasets for foundation models
Domain
Artificial Intelligence / Natural Language Processing (Korean)
Deliverable
production ML models
Required skills
Korean language proficiency, logical reasoning, fact-checking, data cleaning, model evaluation analysis, MS Excel, basic NLP concepts
Preferred skills
Legal/Administrative/Tax background, Legal Tech/AI data project experience, math/science olympiad participation, problem creation experience
Technologies
Foundation models, NLP frameworks
Responsibilities
Generate and review diverse data to enhance model inference, perform data cleaning for logical consistency and factual accuracy, create training and evaluation datasets, analyze model evaluation results to assess performance, handle other Korean NLP related tasks