Overview
Own end-to-end delivery of platform features and services, including technical design, implementation, testing, rollout, and production operation. Build reliable, scalable components for Kubernetes-based infrastructure, improving service deployment, lifecycle management, and operational workflows. Improve engineering quality and developer productivity through automated testing, integration validation, CI/CD systems, and reusable tooling. Design for security, reliability, and operability from the outset, with clear telemetry, actionable diagnostics, controlled rollouts, and recovery mechanisms. Develop automation that reduces repetitive work and operational risk. Where appropriate, incorporate AI-assisted or agent-based approaches with evaluation, access controls, and human oversight. Investigate complex production issues, identify root causes using telemetry, and deliver durable improvem
Full job description
Full Job Description
Own end-to-end delivery of platform features and services, including technical design, implementation, testing, rollout, and production operation. Build reliable, scalable components for Kubernetes-based infrastructure, improving service deployment, lifecycle management, and operational workflows. Improve engineering quality and developer productivity through automated testing, integration validation, CI/CD systems, and reusable tooling. Design for security, reliability, and operability from the outset, with clear telemetry, actionable diagnostics, controlled rollouts, and recovery mechanisms. Develop automation that reduces repetitive work and operational risk. Where appropriate, incorporate AI-assisted or agent-based approaches with evaluation, access controls, and human oversight. Investigate complex production issues, identify root causes using telemetry, and deliver durable improvements to platform reliability and performance. Partner with product managers, engineers, and service owners to clarify requirements, make technical trade-offs, and deliver measurable outcomes. Participate in on-call support and incident response for the services your team owns, translating operational learnings into better software and practices. Mentor engineers through design discussions and code reviews, contributing to a collaborative and inclusive engineering culture. Bachelor's Degree in Computer Science or related technical field AND 4+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python or equivalent experience. These requirements include but are not limited to the following specialized security screenings: Master's Degree in Computer Science or related technical field AND 6+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python OR Bachelor's Degree in Computer Science or related technical field AND 8+ years technical engineering experience with coding in languages including, but not limited to, C, C++, C#, Java, JavaScript, or Python or equivalent experience. Proficiency in one or more programming languages such as Go, C#, or Python. Experience designing, building, and supporting distributed systems, cloud services, or backend infrastructure. Experience taking a significant feature or service from requirements and design through production delivery. Solid coding, debugging, testing, and systems-design skills, with attention to maintainability and operational quality. Ability to communicate technical decisions clearly, collaborate across teams, and work independently through ambiguous problems. Hands-on experience with Kubernetes, containers, and Azure or comparable cloud platforms. Experience with deployment orchestration, control planes, infrastructure automation, or large-scale service lifecycle management. Experience building engineering systems, including CI/CD pipelines, integration tests, release validation, or developer tooling. Experience improving service reliability through observability, incident investigation, performance analysis, and safe deployment practices. Experience applying AI tools or agents to software development or operational workflows, including evaluating their effectiveness and managing their risks.
Tips for this job
Practical JobOpportunity guidance. These tips do not replace official rules or create new eligibility requirements.
- Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
- Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
- Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
- Apply through the original employer or official recruitment destination shown on this page.
Verification notes
Verified from public schema.org JobPosting structured data on the official source page. The complete published description, responsibilities, requirements and benefits were normalized when present; unstated facts were not inferred.
JobOpportunity is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.
Apply through JobOpportunity →Browse current JobOpportunity listings from Microsoft Opportunities →