Skip to main content
Loading
loadingbar
Loading, Please wait..!!

GPU Reliability Engineer - AI Supercomputing Fleet

  • Job type Posted on: Jul 19, 2026
  • Experience level Thinkingmachines
  • Employment type San Francisco, California
  • Remote status Salary: $475,000 per year
  • Employment type Onsite
  • Salary Full-time

Point Apply Here APPLY LATER

Curious about compensation?

Explore the historical salary trends, average pay, and estimated compensation for GPU Reliability Engineer - AI Supercomputing Fleet roles in California.

View Salary Guide →

Job Title :

GPU Reliability Engineer - AI Supercomputing Fleet

Job Type :

Full-time

Job Location :

San Francisco California United States

Remote :

No

Jobcon Logo Job Description :

Thinking Machines in San Francisco is hiring an engineer to ensure the reliability of their GPU supercomputing fleet. You'll be responsible for diagnosing hardware issues and collaborating with vendors to resolve them efficiently. Ideal candidates hold a Bachelor’s degree in computer science or engineering, possess backend programming skills in Python or Rust, and have experience with large-scale systems. The position offers a competitive salary range of $350,000 to $475,000 and generous benefits including health insurance and unlimited PTO. #J-18808-Ljbffr

Jobcon Logo Position Details

Posted:

Jul 19, 2026

Reference Number:

14660_B2C610CF40A002BB2AD4B79F482F0446

Employment:

Full-time

Salary:

Not Available

City:

San Francisco

Job Origin:

APPCAST_CPC

Share this job:

  • linkedin

Jobcon Logo
A job sourcing event
In Dallas Fort Worth
Aug 19, 2017 9am-6pm
All job seekers welcome!

GPU Reliability Engineer - AI Supercomputing Fleet    Apply

Click on the below icons to share this job to Linkedin, Twitter!

Thinking Machines in San Francisco is hiring an engineer to ensure the reliability of their GPU supercomputing fleet. You'll be responsible for diagnosing hardware issues and collaborating with vendors to resolve them efficiently. Ideal candidates hold a Bachelor’s degree in computer science or engineering, possess backend programming skills in Python or Rust, and have experience with large-scale systems. The position offers a competitive salary range of $350,000 to $475,000 and generous benefits including health insurance and unlimited PTO. #J-18808-Ljbffr

Loading
Please wait..!!