Job Description
Lead the design and delivery of multi-quarter initiatives that modernize code flow tools, services, and workflows. Set technical direction for improvements to platform reliability, performance, security, maintainability, and developer experience. Diagnose and troubleshoot complex failures with services and dependencies, driving root-cause fixes rather than recurring mitigations. Partner with product teams and Engineering System organizations to integrate new workflows while protecting compatibility, scale, and operational continuity. Establish measurable success criteria, telemetry, and operational signals that make regressions and customer impact visible. Use and improve AI-assisted engineering workflows for investigation, implementation, review, and operational response while maintaining solid validation and engineering judgment. Mentor engineers through design reviews, code reviews, debugging, and project leadership, raising the technical capability of the broader team. Lead incident response for high impact failures and convert operational learning into systemic reliability improvements. Bachelor's Degree in Computer Science or related technical field AND 4+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python. OR equivalent experience. These requirements include but are not limited to the following specialized security screenings: Master's degree in computer science or a related technical field and 6+ years of technical engineering experience; or Bachelor's degree in computer science or a related technical field and 8+ years of technical engineering experience; or equivalent experience. Experience with developer infrastructure, Azure, CI/CD platforms, or large-scale distributed services. Hands-on experience with Azure DevOps, Azure App Services, GitHub, or similar CI/CD automation. Experience with large-scale telemetry and query systems such as Kusto/KQL, Azure Data Explorer, Application Insights, or similar technologies for diagnosing service and pipeline health. Track record of measurably improving engineering-system throughput, reliability, cost efficiency, or developer productivity for a large engineering organization. Experience mentoring engineers and creating clarity across ambiguous, multi-stakeholder efforts. Experience applying AI-assisted development tools responsibly to improve engineering quality and productivity. Experience designing, building, and operating software used by multiple teams or at significant scale. Demonstrated success leading technically complex projects across team or organizational boundaries. Solid debugging and systems problem-solving skills across unfamiliar codebases and layered infrastructure. Ability to communicate technical strategy, tradeoffs, risks, and decisions clearly to engineering and partner audiences.