Verified current Job

Member of Technical Staff - GPU Infrastructure Engineer

ABOUT LIQUID AI

Job Remote Full source details
Liquid AI San Francisco Source published Sep 20, 2026 Verified 13 hours ago
✓ 100% verification score · Source: Liquid AI (ashby) · Always confirm final requirements on the original source.
Complete source information imported The available role or programme description, requirements, benefits and source facts were imported from the public official endpoint and formatted for reading.
EmploymentFull-time
Work modeRemote / location-flexible

Overview

ABOUT LIQUID AI

Full job description

ABOUT LIQUID AI Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there. THE OPPORTUNITY Our Cluster Infrastructure team owns the compute environments that power foundation model training and research at Liquid AI. We are looking for a hands-on software engineer to keep our GPU clusters reliable, improve resource efficiency, and build the tooling that allows researchers to focus on model development rather than infrastructure. This role matters because infrastructure issues can delay training by days, while improvements in utilization, storage management, and automation can significantly increase research velocity and reduce compute costs. You will work closely with researchers and infrastructure engineers, owning problems from immediate operational response through long-term platform improvements. WHAT WE’RE LOOKING FOR We need someone who:

  • Brings order to complex systems: You identify root causes and build durable fixes rather than repeatedly firefighting.
  • Is an engineer first: You can go deep across Linux, networking, storage, schedulers, and distributed systems.
  • Balances operations and engineering: You handle urgent issues while steadily replacing manual work with automation.
  • Owns outcomes: You communicate clearly, prioritize effectively, and drive problems to resolution across internal teams and external providers. THE WORK
  • Own the reliability and operation of the GPU clusters used for training and research.
  • Debug issues across compute, storage, networking, schedulers, and distributed workloads.
  • Improve CPU, GPU, and storage utilization through better tooling and automation.
  • Onboard and migrate workloads across GPU providers and hardware platforms.
  • Build monitoring, validation, and platform abstractions that reduce operational work for researchers.
  • Contribute to the longer-term architecture of Liquid AI’s training infrastructure and GPU platform. DESIRED EXPERIENCE MUST-HAVE
  • Strong software engineering experience, with the ability to build production-quality infrastructure tooling and automation.
  • Deep knowledge of distributed systems, Linux, networking, and storage.
  • Experience operating a shared compute cluster or distributed training platform.
  • A track record of supporting production users and turning recurring failures into durable solutions.
  • The technical depth to partner effectively with senior research and infrastructure engineers. NICE-TO-HAVE
  • Experience with SLURM, Kubernetes, Ray, Hadoop, or another distributed compute platform.
  • Experience supporting GPU, HPC, or large-scale AI training infrastructure.
  • Experience with distributed storage, cluster schedulers, cloud providers, or infrastructure control planes. WHAT SUCCESS LOOKS LIKE (YEAR ONE)
  1. Researchers spend less time resolving infrastructure and resource-allocation issues.
  2. GPU, CPU, and storage resources are used more efficiently across the fleet.
  3. Recurring operational problems are replaced with automation, monitoring, and dependable platform tooling.
  4. Liquid AI has the beginnings of a durable internal platform that hides infrastructure complexity from researchers. WHAT WE OFFER
  • High-impact ownership: Own infrastructure that directly affects how quickly and efficiently we train foundation models.
  • Compensation: Competitive base salary with equity in a unicorn-stage company.
  • Health: We pay 100% of medical, dental, and vision premiums for employees and dependents.
  • Financial: 401(k) matching up to 4% of base pay.
  • Time Off: Unlimited PTO plus company-wide Refill Days throughout the year.

Tips for this job

Practical Job and Scholarship guidance. These tips do not replace official rules or create new eligibility requirements.

  1. Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
  2. Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
  3. Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
  4. Apply through the original employer or official recruitment destination shown on this page.

Verification notes

laptop-ats-crawler v2

Original authoritative source

Job and Scholarship is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.

Liquid AI (ashby) ↗

Browse current Job and Scholarship listings from Liquid AI (ashby) →

Related opportunities

Other current verified records you may want to review.

Job

Специалист по сопровождению программного обеспечения

Республиканское государственное предприятие на праве хозяйственного ведения Институт ядерной физики Агентства Республики Казахстан по атомной энергии · KZ

Инженер программист Лаборатории Информационных технологий и Искусственного интеллекта

Job

Кафе/мейрамхана әкімшісі

ИСМАГУЛОВ ЕРЛАН ЕРМУХАНОВИЧ · KZ

Еңбекті ұйымдастыру және қызметкерлерді басқару, Кассалық, ұйымдастырушылық-өкімдік, есептік құжаттарды жүргізу, Қонақтарға қызмет көрсету станда...

Job

Автомобиль жүргізушісі

ИП КОМАРОВА · KZ

Әр түрлі машиналарды қауіпсіз басқару, Автомобиль бөлшектерін жөндеу және ауыстыру, Жол қозғалысы ережелерін білу Тәжірибесі жоқ техникалық және...

Job

Бағдарламалық қамсыздандыруды сүйемелдеу жөніндегі маман

Республиканское государственное предприятие на праве хозяйственного ведения Институт ядерной физики Агентства Республики Казахстан по атомной энергии · KZ

Инженер программист Лаборатории Информационных технологий и Искусственного интеллекта

Job

State Boards

Publicjobs Tal Net Opportunities

State Boards is the process through which we select and recommend for appointment, Chairpersons and Members (Non-Executive Directors), to the boa...

Job

Medical Consultants

Publicjobs Tal Net Opportunities

Consultant Psychiatrist Of Learning Disability (Adult) - Mental Health Service Wexford Vacancy type: publicjobs Department/Organisation: HSE Dubl...

More ways to save

Discover deals, coupons and free courses on our sister site.

Explore DealVorio
Save more with DealVorio: deals, coupons, free courses, apps and books