Wafer
Member of Technical Staff
San Francisco
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Lead / management
- Stated salary
- $200,000 – $300,000 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Member of Technical Staff
Inference performance depends on how efficiently models use the underlying hardware. Techniques across kernels, compilers, runtimes, and serving systems can dramatically improve latency and throughput, but these optimizations are difficult, hardware-specific, and slow to reproduce across new accelerators. And none of it counts until it is running in a customer's production traffic.
At Wafer, we are building AI systems that automatically optimize inference workloads across silicon. The goal is fungible token capacity. Any accelerator optimized toward serving inference most efficiently.
Wafer is well funded and serves trillions of tokens a month for mission critical workloads. We serve the highest performance inference to fast-growing AI startups.
Members of Technical Staff build the systems that make that possible and own the customers running on them. There is no separate solutions team, and no layer between you and the workload.
What you'll do
Develop and optimize high-performance computing kernels, and work across inference engine internals and serving infrastructure
Develop AI agents to do autonomous inference engineering
Design, deploy, and operate heterogeneous clusters across vendors
Own customer accounts end to end. Sales gets the first meeting. From there you decide what to prove, you build it, you keep it running in production, and you carry the relationship
Win the technical evaluation. Prove Wafer on the customer's own workload rather than a synthetic benchmark, and be the person who can explain the result to their engineers
Own production for your accounts. When latency moves or error rates climb, you find it, you fix it or route it, and you are who the customer hears from
Inform product roadmap. You sit closer than anyone to how Wafer behaves under real load, and the engineering team builds against what you report
What we look for
You are an exceptional engineer. You have shipped and operated production systems, and you can still open a profiler, read a trace, and find the problem yourself. Inference experience helps, but it matters more that you can be handed an unfamiliar system and own it inside a week.
You have owned a customer relationship. You can name an account that was yours, say what was going wrong, and describe what you personally did about it.
You are credible in a room full of engineers. You can take a skeptical ML team through a benchmark, defend the methodology, and concede the point when they are right.
You work without a spec. The problem arrives half-defined from a customer who does not yet know what they need, and you come back with a scoped answer rather than a list of questions.
How we evaluate
We score every candidate on seven values:
Infinitely Resourceful
Exceptionalism
Unreasonable Standards
Company Over Self
High EQ
Learns Quickly
First Principles Thinker
Compensation and benefits
$200-300K base salary + generous equity.
Fully covered medical, dental, and vision insurance.
Daily lunch and dinner, unlimited PTO, and parental leave.
$1K/month housing stipend (post-tax) if you live within walking distance (0.5 miles) from the office.
Covered Uber/Waymo from/to office.
Visa sponsorship available.
How we work
On-site in San Francisco, five days a week. Small team with massive surface area and ownership. You operate with complete autonomy of how to solve problems. We don't see engineers as code writers, but as problem solvers. You will work across the stack to build a company scaling extremely quickly.
This is not a role that lives in meetings. Nearly all of your time is engineering. What makes it different is who that engineering is for: you are responsible for a named set of customers, their workloads run on what you build, and you are the person they hear from when it matters.
Similar jobs
- Sr. Member Technical Staff - ESD and Latch-Up - HBMMicron · Folsom, CAFirst seen 6d ago
- Member Technical StaffPirros · Los Angeles OfficeFirst seen 31d ago
- Member Technical Staff - Applied AI Engineer (US Timing) Composio · BangaloreFirst seen 26d ago
- Member TechnicalBroadridge · Bengaluru-EPIP Industrial AreaFirst seen 9d ago
- Senior Member TechnicalBroadridge · Hyderabad-Hi-Tec CityFirst seen 11d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job