Job Description
Experience or demonstrated interest in operating system fundamentals is a plus; experience with existing AI inference toolchains is valued. Bachelor's Degree in Computer Science or related technical field AND 4+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR equivalent experience. 5+ years of professional software engineering experience with coding in one or more languages such as C, C++, C#, Java, JavaScript, Python, Rust, or TypeScript, including designing automated tests and validation for the software you ship. Demonstrated knowledge of Server administration and management features. Familiarity with Server deployment and usage strategies in air-gapped or on-premises environments. Understanding of AI model catalogs, model servers, protocols and inference layers. Familiarity with the current state of the art around observability and diagnostics. Experience with systems software development, deep systems architecture, driver development, and code architecture/design for complex software systems. Design and implement how we track usage, policy and efficiency for artificial intelligence capabilities. Build observability pipelines through telemetry insights, diagnostics, and health signals needed to measure how this on-prem infrastructure is performing. Define and develop the management layer for AI inference toolchains from model lifecycles and catalogs to language and protocol gateways. Design and implement validation plans to assess the quality of this infrastructure and evaluate the end to end workflows and life cycle experiences of such a solution. Coordinate with security and compliance teams to ensure this toolchain meets or exceeds platform release requirements.