Senior Site Reliability Developer
Core
Design, develop, and maintain resilient, scalable AWS cloud infrastructure and microservices architecture for Autodesk's exabyte-scale data platform.
Role type
Senior Site Reliability Developer (SRE)
Builds
High-availability cloud data platform components powering desktop, mobile, and web products
Domain
Cloud Infrastructure / DevOps / Data Platform
Deliverable
production ML models | infrastructure
Required skills
AWS resource management, Infrastructure as Code (Terraform, CloudFormation), CI/CD pipeline design (Jenkins, GitHub, Artifactory), Container orchestration (Docker, Kubernetes, ECS), Monitoring and logging (Dynatrace, Grafana, ELK, CloudWatch), Java/SpringBoot, Linux systems administration, Scripting (Python, Go, Bash, Groovy, Node.js)
Preferred skills
AI/ML for DevOps automation, OpenTelemetry, Redis, Data streaming (Kinesis, Firehose, Kafka), JVM optimization, Gradle
Technologies
AWS (ECS Fargate, Lambda, Kinesis, DynamoDB, VPC, IAM, API Gateway, Route 53), Docker, Kubernetes, Jenkins, GitHub, Terraform, Grafana, Dynatrace, ELK Stack, CloudWatch, Java, SpringBoot, Kafka, Flink, Python, Go, Bash, Groovy, Node.js, Redis
Responsibilities
Lead architecture and solution design for cloud microservices; Manage end-to-end implementation and release planning; Streamline CI/CD and automate infrastructure deployment; Implement Disaster Recovery strategies and conduct failover exercises; Provide real-time operational support and participate in on-call rotations; Contribute to CVE remediation and security compliance; Document security best practices across DevOps/SRE pillars.
Seniority
Senior, hands-on IC