Member of Technical Staff (Software Engineer)
Core
Develop and maintain high-performance, low-latency inference infrastructure for AI model deployment and optimization.
Role type
Software Engineer (Inference Infrastructure)
Builds
Scalable inference services, containerized environments, and automated data pipelines for real-time AI tasks.
Domain
Artificial Intelligence / High-Performance Computing / Cloud Infrastructure
Deliverable
production ML models
Required skills
Docker, Kubernetes, Java or C++, Python or Groovy, JavaScript or TypeScript, Linux, SQL, OracleDB, Redis, Git
Preferred skills
ActiveMQ, Kafka
Technologies
Kubernetes, Docker, ActiveMQ, Kafka, Python, Groovy, JavaScript, TypeScript, Linux, SQL, OracleDB, Redis, Git
Responsibilities
Implement infrastructure for high-performance, low-latency inference services; Deploy and configure Kubernetes services; Optimize resource allocation and auto-scaling policies; Integrate inference services with containerized environments; Ensure high availability and fault tolerance via multi-region deployments; Develop Python-based scripts and APIs for data preprocessing and inference execution; Collaborate with ML engineers to validate inference accuracy and performance; Triage and resolve defects by analyzing logs, metrics, and distributed traces; Debug issues related to model deployment, container orchestration, or networking; Develop automated scripts to detect and mitigate failure modes; Author technical documentation for infrastructure configurations and workflows.
Seniority
Entry-level (1+ year experience)