Member of Technical Staff, Applied Research
Core
Develop and train vision-language models for document processing and understanding, building data pipelines and benchmarks to make AI systems more accurate and cost-effective.
Role type
Senior IC applied research engineer (vision-language models)
Builds
Production document AI systems, data pipelines, and evaluation benchmarks
Domain
AI/ML, Document Understanding, Vision-Language Models
Deliverable
production ML models
Required skills
Machine learning engineering, Python, PyTorch, Computer vision, Vision-language models, NLP, Document AI, OCR, Model benchmarking, Synthetic data generation
Preferred skills
Startup experience, Post-training, Fine-tuning, Open-source AI infrastructure, Modern AI coding workflows (vLLM, Pydantic)
Technologies
PyTorch, Python, vLLM, Pydantic
Responsibilities
Develop and train vision-language models for document processing; Build data pipelines for curation, synthetic data generation, and benchmark creation; Evaluate base models and perform post-training or fine-tuning; Improve model accuracy, latency, and cost-effectiveness; Design and maintain benchmarks for extraction quality and system reliability; Collaborate with engineering to move prototypes into production; Work directly with customers to translate requirements into experiments.
Seniority
Senior, hands-on IC