hirly

Jobgether

Senior Software Engineer – LLM Evaluation

Canada

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Jobgether first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.7M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Engineering
Seniority
Senior
Country
CA
Work mode
Remote-friendly
First seen by hirly
6 Oct 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Software Engineer – LLM Evaluation based in Canada.

  • This part-time consulting opportunity is designed for experienced software engineers interested in shaping the evaluation of advanced large language models.
  • You will curate high-quality code, develop software solutions, and assess AI-generated implementations against professional engineering standards.
  • The role spans multiple programming languages and covers the full software-development lifecycle, from architecture and prototyping to deployment, monitoring, and maintenance.
  • You will also help create automated verification mechanisms and benchmarks that make model evaluation more rigorous and reproducible.
  • Working alongside research and technical teams, you will help identify model strengths, weaknesses, and recurring coding errors.
  • Your engineering judgment will directly contribute to improving coding-focused AI evaluation systems and research.
  • The engagement is fully remote and flexible, with a minimum commitment of 10 hours per week and the potential to work up to 40 hours weekly.
Original posting on Jobgether's site ↗

Listed on hirly, a job board. hirly is not the employer: Jobgether is hiring for this role.

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job