hirly

Clera

Staff Engineer, Agentic AI

San Francisco

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Clera first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

hirly's read of this role

Role family
Engineering
Seniority
Lead / management
Stated salary
$160,000 – $250,000 per year
Country
US
Work mode
On-site / unstated
First seen by hirly
30 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

About the Role

This is a senior technical leadership role at the heart of an early-stage AI software company building intelligent agents for hardware engineers. You will own the core agent intelligence layer that turns engineers' intent into reliable, cost-efficient multi-step workflows across desktop CAD, simulation, and PLM tools. Reporting directly to the CTO, the work you do here determines the product's real-world value to enterprise customers.

What You'll Do

Lead development of the agent intelligence layer that executes multi-step workflows across complex desktop engineering software.

Serve as technical lead for a small team of AI engineers, a user researcher, and domain expert contractors.

Own the full product loop: define agent capabilities from user stories, build implementations, and benchmark against real workflows.

Drive agent task success rate by defining evaluation frameworks, establishing baselines, and iterating on completion metrics.

Set and enforce per-task token budgets and track cost per completed workflow to ensure commercial viability.

Build rigorous, reproducible evaluation infrastructure grounded in validated user stories.

Lead user story mapping and validation through interviews and close collaboration with domain experts.

Translate validated user stories into testable evals, closing the loop between user research and benchmarking.

Own agent architecture decisions including tool-calling, state management, error recovery, model routing, and context management.

Act as a player-coach: write production code, review designs, unblock the team, and raise engineering standards.

Collaborate cross-functionally with integrations, product, and customers during POCs to align agent behavior with real-world usage.

What We're Looking For

7+ years of software engineering experience, including at least 2 years building LLM-based agents that take real-world actions.

Deep experience designing LLM application architectures: model selection, context and window management, retrieval, tool calling, and orchestration patterns.

Hands-on experience shipping AI or LLM tooling on top of proprietary engineering data or desktop engineering software (for example, agents or MCP servers over CAD, PLM, or simulation platforms); general-purpose chatbot or web-app RAG work alone does not qualify.

Strong Python proficiency and familiarity with LLM function calling, tool APIs, observability and tracing, and evaluation frameworks.

Proven ability to build evaluation and benchmarking frameworks measuring task completion, cost efficiency, and failure modes.

Technical leadership experience: setting direction for small teams of 3 to 6 engineers and performing meaningful code review while continuing to write production code.

Experience with desktop automation or programmatic control of applications such as COM or similar interfaces.

Domain background in mechanical engineering, CAD, CAE, PLM, or an adjacent engineering-software field.

Familiarity with enterprise deployment constraints on locked-down corporate workstations.

Track record contributing to public benchmarks, publications, or open-source agentic AI projects is a plus.

Compensation & Benefits

Base salary range: $160,000 to $250,000 USD annually , plus equity. Visa sponsorship is not available for this role.

Location

On-site in San Francisco, California, United States .

Original posting on Clera's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job