hirly

NexGen Cloud

Infrastructure Operations Engineer

Quebec, Canada

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at NexGen Cloud first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Seniority
Mid level
Country
CA
Work mode
Remote-friendly
First seen by hirly
3 Oct 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

ABOUT NEXGEN CLOUD:

NexGen Cloud is the company behind Hyperstack, a full-stack AI cloud serving tens of thousands of customers from AI researchers to enterprises running the world's most compute-intensive workloads. We deliver on-demand and private GPU infrastructure to teams who treat performance as a requirement, not a feature.

We're a tight-knit, fast-moving team working at the cutting edge of AI cloud infrastructure. We practice what we preach, equipping our people with AI at every level so we can solve harder problems, ship faster, and keep raising the bar for what enterprise GPU infrastructure looks like.

THE ROLE: Infrastructure Operations Engineer

This role exists because our platform is scaling quickly — and complexity comes with it. As we expand our OpenStack and Kubernetes environments globally, we need engineers who can take real ownership of how the platform is designed, operated, and improved. You'll have direct ownership over business-critical infrastructure that impacts performance, reliability, and customer experience.

This is not a maintenance role. If you like solving hard problems, owning systems end-to-end, and seeing the impact of your work immediately — you'll enjoy this.

WHAT YOU'LL BE DOING:

Rather than a long checklist, here's what success in this role looks like:

Own the design, deployment, and operation of OpenStack and Kubernetes environments — ensuring platform performance, scalability, and resilience for GPU workloads

Build and improve infrastructure using infrastructure-as-code and GitOps practices, driving automation across provisioning, deployment, and operational workflows

Optimise GPU workload scheduling using Kubernetes and NVIDIA tooling, and implement monitoring, logging, and alerting to ensure platform stability

Lead incident response and drive continuous improvement of reliability across the platform

Maintain strong security controls across infrastructure and container layers — RBAC, network policies, and tenant isolation

Work closely with Platform, DevOps, AI, Product, and Support teams to align infrastructure capabilities with customer and platform requirements

ABOUT YOU:

We're more interested in how you think and work than in a perfect CV. You'll likely bring a combination of the following:

Essential

Extensive hands-on Linux systems administration skills and knowledge — genuine depth, not surface-level familiarity

Strong, proven experience building servers and racks — you've physically assembled, cabled, and commissioned hardware, not just specified or overseen it

Direct hands-on experience physically working in data centres — you've personally stacked and racked hardware on-site

A willingness and ability to travel to Quebec sites as required

A solid understanding of networking and storage systems

Nice to Have

Experience installing, racking, and configuring GPU hardware specifically, ideally including NVIDIA platforms

Production experience running OpenStack and/or Kubernetes at scale

Experience with infrastructure automation, CI/CD, and Git-based workflows

Broader exposure to HPC or large-scale compute environments

Contributions to open-source projects

WHAT WE OFFER:

Competitive salary and annual discretionary bonus scheme

Employee wellbeing benefits

25 days of holiday, plus public holidays

Flexible working arrangements (remote or hybrid, depending on role and location)

Real ownership and autonomy, with the trust to take initiative and experiment

The opportunity to make a visible, meaningful impact as we scale

Clear career progression and growth opportunities in a fast-growing company

A collaborative, international culture built on trust, transparency, and ownership

The chance to help shape NexGen Cloud's team, culture, and future alongside ambitious, mission-driven colleagues

MORE INFORMATION

Head over to our NexGen Cloud careers page to view current openings and follow us on LinkedIn and X to learn more about our journey, newest releases and hear exciting news in the neocloud space.

Original posting on NexGen Cloud's site ↗

Listed on hirly, a job board. hirly is not the employer: NexGen Cloud is hiring for this role.

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job