Job Requisition ID #
26WD100437
26WD100437 Senior DevOps Developer
L'affichage de poste en français suivra / The French job posting follows
Position Overview
At Autodesk, we are changing how people design and create, empowering them to imagine, design, and make a better world. We believe in the power of automation and cutting-edge technology to inspire creativity. The Bidding group within Autodesk Construction Solutions is looking for an experienced Site Reliability Engineer (SRE)/DevOps Developer to help build, operate, and continuously improve highly available, resilient, and scalable cloud-native services that power our customers' critical workflows.
In this role, you will partner closely with Software Engineering, Infrastructure, and Platform teams to improve developer productivity, strengthen platform reliability, automate operational processes, and enhance observability across production services. You will own service health, operational excellence, disaster recovery readiness, security compliance, and CI/CD automation while driving continuous improvements in monitoring, infrastructure as code, and incident management.
Your work will include building and maintaining Kubernetes-based workloads, AWS infrastructure, MongoDB-backed services, observability platforms such as Dynatrace, cloud automation, disaster recovery solutions, and secure operational processes that enable engineering teams to deliver software reliably and efficiently.
You will report to the Manager, Software Development as part of a remote team (distributed engineering team).
Minimum Responsibilities
Collaborate with Infrastructure & Systems teams to design and implement innovative solutions that improve developer velocity, infrastructure resiliency, security, and availability
Establish, hone, and facilitate adherence to team SLO / SLI targets
Drive processes, including service reviews, fire drills, chaos testing, incident response, and incident post mortem's
Collaborate with partners to understand requirements, use cases, and build towards a cohesive strategy
Develop scripts and tooling to execute deployment and configuration for service updates
Engage in deep technical discussions that shape robust, performant solutions
Coordinate and perform major upgrades with zero downtime
Keep systems updated for security compliance
Troubleshoot complex technical problems
Solve live performance, stability issues and prevent recurrence
Participate in on-call rotations to support production systems