hirly

Wafer

Member of Technical Staff

San Francisco

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Wafer first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Seniority
Lead / management
Stated salary
$200,000 – $300,000 per year
Country
US
Work mode
On-site / unstated
First seen by hirly
28 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Member of Technical Staff

Inference performance depends on how efficiently models use the underlying hardware. Techniques across kernels, compilers, runtimes, and serving systems can dramatically improve latency and throughput, but these optimizations are difficult, hardware-specific, and slow to reproduce across new accelerators. And none of it counts until it is running in a customer's production traffic.

At Wafer, we are building AI systems that automatically optimize inference workloads across silicon. The goal is fungible token capacity. Any accelerator optimized toward serving inference most efficiently.

Wafer is well funded and serves trillions of tokens a month for mission critical workloads. We serve the highest performance inference to fast-growing AI startups.

Members of Technical Staff build the systems that make that possible and own the customers running on them. There is no separate solutions team, and no layer between you and the workload.

What you'll do

Develop and optimize high-performance computing kernels, and work across inference engine internals and serving infrastructure

Develop AI agents to do autonomous inference engineering

Design, deploy, and operate heterogeneous clusters across vendors

Own customer accounts end to end. Sales gets the first meeting. From there you decide what to prove, you build it, you keep it running in production, and you carry the relationship

Win the technical evaluation. Prove Wafer on the customer's own workload rather than a synthetic benchmark, and be the person who can explain the result to their engineers

Own production for your accounts. When latency moves or error rates climb, you find it, you fix it or route it, and you are who the customer hears from

Inform product roadmap. You sit closer than anyone to how Wafer behaves under real load, and the engineering team builds against what you report

What we look for

You are an exceptional engineer. You have shipped and operated production systems, and you can still open a profiler, read a trace, and find the problem yourself. Inference experience helps, but it matters more that you can be handed an unfamiliar system and own it inside a week.

You have owned a customer relationship. You can name an account that was yours, say what was going wrong, and describe what you personally did about it.

You are credible in a room full of engineers. You can take a skeptical ML team through a benchmark, defend the methodology, and concede the point when they are right.

You work without a spec. The problem arrives half-defined from a customer who does not yet know what they need, and you come back with a scoped answer rather than a list of questions.

How we evaluate

We score every candidate on seven values:

Infinitely Resourceful

Exceptionalism

Unreasonable Standards

Company Over Self

High EQ

Learns Quickly

First Principles Thinker

Compensation and benefits

$200-300K base salary + generous equity.

Fully covered medical, dental, and vision insurance.

Daily lunch and dinner, unlimited PTO, and parental leave.

$1K/month housing stipend (post-tax) if you live within walking distance (0.5 miles) from the office.

Covered Uber/Waymo from/to office.

Visa sponsorship available.

How we work

On-site in San Francisco, five days a week. Small team with massive surface area and ownership. You operate with complete autonomy of how to solve problems. We don't see engineers as code writers, but as problem solvers. You will work across the stack to build a company scaling extremely quickly.

This is not a role that lives in meetings. Nearly all of your time is engineering. What makes it different is who that engineering is for: you are responsible for a named set of customers, their workloads run on what you build, and you are the person they hear from when it matters.

Original posting on Wafer's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job
Member of Technical Staff – Wafer · San Francisco | hirly.me