Full Stack Engineer, Fleet Scheduling
Core
Design, develop, and operate web-based systems providing intuitive interfaces for monitoring, scheduling, and managing large-scale AI workloads on supercomputing clusters.
Role type
Senior Full Stack Engineer (AI Infrastructure)
Builds
Web applications for real-time job scheduling, resource tracking, and monitoring of exascale AI clusters.
Domain
AI Infrastructure / Supercomputing
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Full-stack development, modern frontend frameworks (React, Vue, Angular), backend technologies (Python, Go, Node.js), RESTful and GraphQL APIs, distributed databases, cloud infrastructure (Azure), high-performance web application development, distributed systems architecture.
Preferred skills
Kubernetes, Docker, cloud-native deployment, AI/ML workload scheduling and orchestration, real-time data processing, visualization libraries, observability tooling.
Technologies
React, Vue, Angular, Python, Go, Node.js, Azure, Kubernetes, Docker, GraphQL, REST APIs
Responsibilities
Design and develop full-stack web applications to track and manage large-scale AI workloads; Build data visualization tools (Gantt charts, dashboards) for job scheduling insights; Optimize backend services for massive data throughput and low-latency performance; Implement frontend components for seamless interaction with compute systems; Collaborate with researchers and infrastructure teams to translate operational needs into scalable solutions; Ensure system security, reliability, and scalability across distributed infrastructure.
Seniority
Senior, hands-on IC