Job Description
* Own the design and delivery of complex software components and services within Azure infrastructure, driving work from requirements through production. Collaborate with stakeholders to clarify requirements, identify dependencies, evaluate trade-offs, and develop scalable technical designs. Design, implement, optimize, and maintain secure, reliable, and scalable solutions for complex distributed-system scenarios. Use telemetry, diagnostics, testing, and operational insights to improve reliability and prevent production incidents. Lead incident response and retrospectives, driving systemic corrective actions across components and services. Provide technical leadership through design reviews, high-quality code reviews, mentoring, and guidance for other engineers. Drive improvements in engineering systems, automation, deployment safety, observability, and operational excellence. Act as a Designated Responsible Individual (DRI) during on-call rotations, coordinating responses to complex production issues, communicating status to stakeholders, restoring service, and driving follow-up improvements. Bachelor's Degree in Computer Science or a related technical field AND 5+ years of technical engineering experience coding in languages including, but not limited to, C#, C++, Go, Java, Kotlin, or Python. Strong backend software engineering skills, including designing, implementing, testing, and operating production services. Experience designing and developing scalable distributed systems. Strong software design, debugging, and problem-solving skills. Bachelor's Degree in Computer Science or a related technical field AND 8+ years of technical engineering experience coding in languages including, but not limited to, C#, C++, Go, Java, Kotlin, or Python. OR Master's Degree in Computer Science or a related technical field AND 6+ years of technical engineering experience coding in those languages. OR equivalent experience. Experience developing distributed systems and services in scalable cloud environments. Experience owning and operating production services, including monitoring, incident response, reliability improvements, and operational readiness. Experience with datacenter infrastructure, hardware lifecycle management, or large-scale resource orchestration. Experience providing technical leadership through architecture and design reviews, code reviews, and mentoring engineers. Demonstrated commitment to software quality, security, performance, maintainability, automation, and operational excellence. These requirements include, but are not limited to, the following specialized security screening: