hirly

AfterQuery

Research Scientist – Frontier Evaluations

San Francisco

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at AfterQuery first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.6M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Mid level
Stated salary
$210,000 – $450,000 per year
Country
US
Work mode
On-site / unstated
First seen by hirly
28 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About AfterQuery

AfterQuery is an applied research lab curating data solutions for foundation model development. We serve every frontier AI lab with the mission of delivering the best data to power the best models. In doing so, we can make expertise that once took a lifetime to build available to anyone who needs it.

Our customers are the ones building the foundation models themselves and our work sits directly in the loop of how those systems improve. This is a rare opportunity to join a company at a defining moment in AI. Forbes reported that we could be YC's fastest unicorn, reportedly raising at a $3.2 billion valuation. We're based in San Francisco and backed by leading investors including Altos Ventures, BoxGroup, and Y Combinator and angels from Google DeepMind, OpenAI, Anthropic, Meta Superintelligence Labs, and Microsoft AI.

Why Apply

Massive Opportunity: Forbes reported that we could be YC's fastest unicorn, reportedly raising at a $3.2 billion valuation, and we're not slowing down.

Founding Impact: You will own and architect core infrastructure systems that power our platform from the ground up.

Equity & Growth: Competitive salary and meaningful equity. As we scale, you’ll have the opportunity to shape the engineering organization and lead major technical initiatives.

Strong Team: Our founding team has experience from Citadel Securities, Meta, Google, Silver Lake, and Morgan Stanley — work alongside world-class engineers and researchers.

Overview

AfterQuery is hiring Research Scientists to design and publish rigorous evaluations for frontier AI systems. The role spans agentic, coding, and safety evaluations, as well as expert-domain evaluations involving applied AI in healthcare, STEM, finance, and related fields. You will own evaluation development end to end and collaborate across disciplines to turn important capability gaps into rigorous public research.

Responsibilities

Lead the end-to-end design, validation, launch, and continuous improvement of frontier AI benchmarks.

Partner with researchers and domain experts to develop evaluations around meaningful model failures, gaps in existing coverage, and high-priority domains.

Analyze model capabilities and failure modes using rigorous experimental design and statistical methods.

Build reproducible evaluation systems, including harnesses, graders, and benchmark infrastructure.

Collaborate with researchers to post-train models and measure the resulting performance gains.

Communicate results through benchmark reports, technical articles, and research papers.

Required Qualifications

Strong record of publishing benchmarks or research papers.

Clear technical communication and strong scientific writing skills.

Commitment to experimental rigor, including baselines, ablations, statistical validity, and contamination controls.

Ability to take an ambiguous evaluation question from initial scoping through a reproducible public release.

Depth in agentic, coding, and safety evaluations or applied machine learning in an expert domain.

Preferred Qualifications

PhD in a related technical field.

Research publications at leading conferences or peer-reviewed journals.

Interest in multidisciplinary research and the creativity to combine methods and insights from AI, engineering, science, and other expert domains.

Company Benefits (For Eligible Employees):

Health Insurance: Medical, Vision, Dental

401(k) with Employer Match

Daily Meals: Daily UberEats Stipend

Monthly Wellness Stipend

We are an equal opportunity employer committed to providing a workplace free from discrimination and harassment. Employment decisions are made without regard to legally protected characteristics under applicable federal, state, or local law.

We comply with applicable pay transparency requirements and provide compensation ranges based on the position, qualifications, experience, and other relevant factors. Reasonable accommodations are available to qualified individuals with disabilities and for sincerely held religious beliefs, as required by law. This job description is intended to describe the general nature and level of work performed and is not an exhaustive list of all duties, responsibilities, qualifications, or working conditions associated with the position. We reserve the right to modify this job description as business needs change.

Original posting on AfterQuery's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job
Research Scientist – AfterQuery · San Francisco | hirly.me