Staff Software Engineer - OCR / Text Extraction
Core
Design, architect, and deliver scalable, event-driven platforms for document content extraction and transformation, integrating OCR technologies and AI services for legal professionals.
Role type
Staff Full Stack Software Engineer (OCR/Text Extraction)
Builds
Production-grade document management systems, semantic search capabilities, and event-driven data pipelines on AWS.
Domain
Legal Technology / Document Management / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
C#/.NET, OCR technologies (Tesseract, Apryse), event-driven architecture, microservices, AWS, Kafka, observability (logging, metrics, tracing)
Preferred skills
AI-driven services integration, semantic search implementation, technical leadership, mentoring
Technologies
AWS, Kafka, Tesseract, Apryse, C#/.NET, DataDog
Responsibilities
Set technical direction for content extraction teams; lead architectural decisions for OCR integration; design and implement scalable APIs; mentor engineers; build resilient systems; optimize performance and cost.
Seniority
Staff, hands-on IC with technical leadership