Verified current Job

Member of Technical Staff - Datacenter Networking

OWN YOUR INTELLIGENCE

Job Remote Full source details
Prime Intellect San Francisco, San Francisco, Remote Source published Sep 20, 2026 Verified 10 hours ago
✓ 100% verification score · Source: Prime Intellect (ashby) · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.
Work modeRemote / location-flexible

Overview

OWN YOUR INTELLIGENCE

Full job description

OWN YOUR INTELLIGENCE Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team. Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own. Prime Intellect has raised $150M in total funding from Founders Fund, Radical Ventures, NVIDIA, and exceptional AI, infrastructure, and enterprise operators — including Andrej Karpathy, Dwarkesh Patel, and leaders and founders from Ramp, Perplexity, Harvey, Mercor, Zapier, Datadog, Cognition, OpenAI, Thinking Machines, Together AI, SemiAnalysis, LangChain, Browserbase, Cloudflare, Sierra, Databricks, Airbnb, OpenRouter, Standard Intelligence, Fleet, Core Auto, and more. We are looking for people who want to build at the intersection of frontier research, real infrastructure, and go-to-market for a category that does not fully exist yet. ROLE IMPACT You'll design and operate the networks that connect large GPU clusters. Own the reliability and performance of training fabrics, storage networks, and management connectivity so distributed workloads can scale without the network becoming the bottleneck. CORE TECHNICAL RESPONSIBILITIES

  • Design and deploy scalable datacenter network topologies for GPU training, inference, storage, and management traffic
  • Configure and operate high-performance Ethernet/RoCE and InfiniBand fabrics with clear standards for routing, redundancy, and capacity
  • Automate network provisioning, configuration validation, upgrades, and rollback procedures
  • Diagnose packet loss, congestion, link failures, and collective communication performance across hosts and switches
  • Benchmark end-to-end network performance with infrastructure and ML teams, translating workload needs into measurable acceptance criteria
  • Build monitoring for port health, errors, utilization, congestion, and fabric topology; improve incident response and runbooks
  • Partner with datacenter operators and hardware vendors on cabling, optics, deployment readiness, and failure resolution TECHNICAL REQUIREMENTS REQUIRED EXPERIENCE
  • 3+ years of production datacenter networking experience
  • Strong understanding of Ethernet, TCP/IP, routing, switching, and redundant network design
  • Hands-on experience with high-performance GPU networking using InfiniBand or RoCE
  • Experience troubleshooting network problems across Linux hosts, NICs, switches, and physical links
  • Ability to automate network operations with Python, Ansible, or comparable tools INFRASTRUCTURE SKILLS
  • Leaf-spine architectures, BGP, ECMP, VLANs, and network segmentation
  • RDMA concepts and performance tuning; congestion control and lossless Ethernet considerations
  • Linux networking, NIC drivers and firmware, packet capture, and throughput/latency testing
  • Optics, transceivers, cable management, and link-level diagnostics
  • Safe change management, configuration versioning, telemetry, and alerting NICE TO HAVE
  • Experience operating 400G/800G networks or large multi-rack GPU clusters
  • NVIDIA Spectrum or Quantum networking experience
  • NCCL performance analysis and distributed training troubleshooting
  • EVPN/VXLAN, SONiC, or network source-of-truth systems
  • Experience with network simulation, automated validation, and capacity planning GROWTH OPPORTUNITY You'll work directly with customers pushing the boundaries of AI, from startups training foundation models to enterprises deploying massive inference infrastructure. You'll collaborate with our world-class engineering team while having direct impact on systems powering the next generation of AI breakthroughs. We value expertise and customer obsession - if you're passionate about building reliable, high-performance GPU infrastructure and have a track record of successful large-scale deployments, we want to talk to you. Apply now and join us in our mission to democratize access to planetary scale computing. COMPENSATION Cash compensation range of $150,000–$300,000 plus equity incentives.

Tips for this job

Practical Job and Scholarship guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

laptop-ats-crawler v2

Original authoritative source

Job and Scholarship is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Prime Intellect (ashby) ↗

Browse current Job and Scholarship listings from Prime Intellect (ashby) →

Related opportunities

Other current verified records you may want to review.

Job

Referendariat Wahlstation Arbeitsrecht

Bauindustrieverband Ost Careers

Current Referendariat Wahlstation Arbeitsrecht opening at Bauindustrieverband Ost Careers in Magdeburg. Full employer-published role sections hav...

Job

Buyer II - AMZ24149.3

Amazon.com Services LLC - A57 · United States

Employer: Amazon.com Services LLC Position: Buyer II - AMZ24149.3 Location: Atlanta, GA Multiple Positions Available: Delivering improved financi...

Job

Software Development Engineer, Agentic AI, Velocity Labs

Amazon Development Center U.S., Inc. · United States

The Velocity Labs team mission is to think beyond the confines of the normal product-orientated approach and to discover new ways to apply and em...

Job

Senior Marketing Manager, Global Executive Marketing, AWS Global Executive Marketing

Amazon Web Services, Inc. · United States

AWS Global Executive Marketing (GEM) is responsible for how AWS engages its most senior customer leaders. We design and deliver the strategy, pro...

Job

Sr. Product Manager - Tech, Devices and Services FinTech

Amazon.com Services LLC · United States

Are you interested in working with the teams that developed the Kindle, FireTV and Alexa? The Amazon Devices and Service Finance organization is...

Job

Software Development Manager, Agentic AI, Velocity Labs

Amazon Development Center U.S., Inc. · United States

The Velocity Labs team mission is to think beyond the confines of the normal product-orientated approach and to discover new ways to apply and em...

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books