About the job
Core
Generate, curate, and review data to improve model inference capabilities and evaluate foundation model performance.
Role type
Data generation and evaluation specialist for NLP foundation models
Builds
Training datasets and evaluation datasets for foundation models
Domain
Natural Language Processing (NLP) and Legal/Administrative domains
Deliverable
production ML models
Required skills
Data generation, data curation, logical consistency review, fact-checking, accuracy verification, model performance analysis, Korean language proficiency, NLP background knowledge
Preferred skills
Legal/Administrative/Tax experience in government agencies, Legal Tech or AI data project experience, Math Olympiad experience, Science Olympiad experience, algorithm competition experience, math/science teaching or problem-setting experience
Technologies
None explicitly stated
Responsibilities
Generate diverse data to enhance model inference, review acquired data, perform curation tasks (logical completeness, factual accuracy, precision), create learning and evaluation data, analyze model evaluation results
Seniority
Junior to Mid-level, hands-on IC