Verified current Job

Staff Cloud Native Software Engineer

About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by...

Job Full source details
Nscale (greenhouse) Verified 7 hours ago
✓ 80% verification score · Source: Nscale (greenhouse) · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.

Overview

About Nscale Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility. We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future. About the Role We're hiring a Staff Cloud Native Software Engi

Full job description

Full Job Description

About Nscale

Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.

We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.

About the Role

We're hiring a Staff Cloud Native Software Engineer to build, operate, and improve the cloud-native software integrations that connect AI applications and networking components at scale.

In this software engineering role, you'll work on shared Kubernetes-based platforms, deployment patterns, observability foundations, infrastructure architecture, and operational tooling that help internal teams run services safely and efficiently on GPU-backed infrastructure. You'll partner closely with platform engineering, infrastructure, and product teams to ensure capabilities meet real developer and operational needs.

This role is important to the reliability, scalability, and usability of Nscale's software integrations. As a Staff engineer, you'll take ownership of significant components and set technical direction across teams, deliver complex technical work independently, and raise the quality of operations and engineering through practical improvements, sound technical judgement, and mentoring.

What You’ll Do

Cloud Native Software Engineering

  • Design, build, and operate Kubernetes-native software — controllers, operators, custom resources (CRDs), and admission webhooks — that connects AI applications with core networking components on GPU-backed infrastructure.

  • Extend Kubernetes control-plane capabilities to support AI workload requirements, including network policy controllers, CNI/service-mesh integrations, and resource/scheduling extensions.

  • Own significant components end-to-end and set the technical direction for how they're designed, deployed, and operated across the team.

  • Build reconciliation loops, informers, and client-go–based tooling that keep infrastructure state consistent between the API server, networking systems, and AI runtime components.

  • Develop operational tooling and automation that make Kubernetes-native services easier for internal teams to deploy, run, and support.

Infrastructure Architecture, Reliability & Observability

  • Drive infrastructure architecture decisions around how AI applications and networking components integrate across the platform, weighing trade-offs at a cross-team level.

  • Build observability foundations for controller and operator software — metrics, structured events, tracing, and status reporting surfaced through the Kubernetes API and platform dashboards.

  • Design systems that degrade gracefully and self-heal, using controller patterns (reconciliation, backoff, status conditions) to reduce manual intervention.

  • Debug and resolve complex issues spanning the Kubernetes control plane, networking (CNI, service mesh, kube-proxy/eBPF datapaths), and workload runtime behavior on GPU-backed infrastructure.

  • Define standards for safe rollout of controller and platform changes, including versioning, compatibility, and staged deployment.

Team Technical Leadership

  • Set technical direction for how the team builds Kubernetes-native software, establishing patterns for controller design, CRD schema evolution, and testing strategy.

  • Lead design discussions and code reviews, holding a high bar for Kubernetes API conventions and idiomatic client-go usage.

  • Partner with platform engineering, infrastructure, and product teams to translate real developer and operational needs into clean CRDs, APIs, and controller-managed abstractions.

  • Define reusable patterns, shared libraries, and scaffolding that let other teams build correctly on the platform without reinventing integration logic.

  • Mentor engineers in Kubernetes internals, controller-runtime patterns, and sound operational judgement.

KPIs

  • Reliability, scalability, and usability of AI infrastructure–networking software integrations

  • Correctness and maintainability of Kubernetes controllers and operators in production

  • Reduction in manual operational effort and config drift across supported components

  • Adoption of shared patterns/frameworks and effectiveness of observability tooling across teams

Who You are

  • At least 8 years of experience in production-level software development.

  • Deep hands-on experience building and operating Kubernetes-native software: custom controllers, operators, CRDs, or admission webhooks — using controller-runtime, client-go, or equivalent.

  • Strong understanding of Kubernetes internals: the API server, informer/lister patterns, reconciliation loops, and the object model.

  • Strong networking fundamentals — CNI, service mesh, kube-proxy/eBPF datapaths, DNS, load balancing — and experience building software that integrates with these systems.

  • Proficiency in Go (strongly preferred) or a similar language, with a track record of shipping well-tested, production-quality code at scale.

  • Experience with observability practices — metrics, tracing, structured logging — built into software rather than added afterward.

  • Comfortable owning components independently end-to-end, from design through operation, while setting direction for adjacent teams.

  • Experience with or strong interest in GPU-backed infrastructure and AI workload patterns is a plus.

  • Track record of leading technical design at a staff level and mentoring engineers through practical technical guidance.

What We Can Offer You

You’ll have the opportunity to help shape the operating standards behind a next-generation AI cloud platform, working on complex infrastructure challenges with real ownership and impact. This is a chance to play a meaningful role in scaling high-performance, sustainable data centre operations in a fast-moving environment.

Equal Opportunities Statement

We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.

If there’s anything we can do to accommodate your specific situation, please let us know.

The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.

For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice: Here.

Nscale does not accept unsolicited candidate submissions from recruitment agencies.

Tips for this job

Practical Job and Scholarship guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

Discovered directly from the employer’s public Greenhouse Job Board API with content=true. The complete public job-ad body was normalized into safe candidate-facing content.

Original authoritative source

Job and Scholarship is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Nscale (greenhouse) ↗

Browse current Job and Scholarship listings from Nscale (greenhouse) →

Related opportunities

Other current verified records you may want to review.

Job

Werkstudent Go-To-Market & Partner Management (m/w/d)

Bees Bears Gmbh Careers

Current Werkstudent Go-To-Market & Partner Management (m/w/d) opening at Bees Bears Gmbh Careers in Berlin. Full employer-published role sections...

Job

Senior Mechanical Engineer

Anduril Industries

Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By...

Job

Electrical Technician

Anduril Industries

Anduril Industries is a defense technology company with a mission to transform U.S. and allied military capabilities with advanced technology. By...

Job

Администратор серверов

Республиканское государственное предприятие на праве хозяйственного ведения Институт ядерной физики Агентства Республики Казахстан по атомной энергии · KZ

Администрирование систем защиты информации ИС, Администрирование систем шифрования данных, Восстановление параметров ПО сетевых устройств Без опы...

Job

Сатушы-кассир

ИП Савченко Э · KZ

Тауар өнімін қабылдау және тауарларға баға белгілеуді рәсімдеу, Сатылатын тауарлардың сапасы мен санын тексеру, Сату жоспарын орындау, Орналастыр...

Job

Директордың тәрбие ісі жөніндегі орынбасары

Коммунальное государственное учреждение Средняя школа имени Сакена Сейфуллина отдела образования Сарысуского района управления образования акимата Жамбылской области · KZ

Тәрбие іс-шараларын өткізу, Семинарларды ұйымдастыру және өткізу, Оқу-тәрбие процесін жоспарлау және ұйымдастыру Тәжірибесі жоқ жоғары

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books