hirly

Vast.ai

GPU Systems Engineer – HPC / Parallel Computing

San Francisco · Los Angeles

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Vast.ai first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Seniority
Mid level
Stated salary
$200,000 – $330,000 per year
Country
US
Work mode
On-site / unstated
First seen by hirly
28 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About Us

Vast.ai ’s cloud powers AI projects and businesses all over the world. We are democratizing and decentralizing AI computing—reshaping our future for the benefit of humanity.

We are a growing and highly motivated team dedicated to an ambitious technical plan. Our structure is flat, our ambitions are out‑sized, and leadership is earned by shipping excellence.

We seek engineers with strong intrinsic drive, a true passion for advancing the state of the art, and a mix of architecture, coding, and communication skills.

LOCATION: On-site at our office in San Francisco or Westwood, Los Angeles.

About the Role

We’re looking for a systems engineer with HPC or parallel programming experience to help scale AI inference. You’ll leverage your knowledge of high-performance systems to optimize GPU performance at the bleeding edge of AI.

Full-Time

On-site at either our SF or LA offices

Tech Stack

CUDA/C++, GPGPU, Python, Linux

Key Responsibilities

Design and optimize GPU kernels and tensor libraries

Translate HPC techniques into scalable AI inference solutions

Evaluate emerging architectures and resource management approaches

Collaborate with technical leadership to improve GPU infrastructure efficiency

Ideal Experience

Advanced C++ (C++17/20 preferred)

Expertise with at least one parallel framework (CUDA, HIP, SYCL, OpenCL, OpenACC, or similar)

Strong background in systems optimization and HPC performance tooling

Familiarity with distributed training/inference frameworks (bonus)

Interview Process

After submitting your application, our technical team reviews your credentials. If selected, you'll proceed through the following stages:

15 min - Initial screening (virtual)

45 min - Quick dive into Vast, work history (virtual)

45 min - Systems and architectures (virtual)

1 hour - LLM-assisted coding assessment (virtual)

2 hours - Meet and greet with coding assessment (on-site)

Our goal is to complete the interview process in two weeks.

Benefits

Comprehensive health, dental, vision, and life insurance

401(k) with company match

Meaningful early-stage equity

Onsite meals, snacks, and close collaboration with founders/tech leaders

Ambitious, fast-paced startup culture where initiative is rewarded

Original posting on Vast.ai's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job
GPU Systems Engineer – Vast.ai · San Francisco | hirly.me