hirly

Alkami

Sr. Platform Engineer

US Remote

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Alkami first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Engineering
Seniority
Senior
Country
US
Work mode
Remote-friendly
First seen by hirly
27 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Alkami is the digital sales and service platform provider for U.S. banks and credit unions. Our unified Platform integrates onboarding, digital banking, and data and marketing—each solution can stand alone, but together they deliver more—to help institutions onboard, engage, and grow relationships. As the future shifts toward Anticipatory Banking, we help data-informed bankers meet the moment with technology that drives action.

Founded in 2009, we continue to be recognized for our intentional culture and tremendous growth (Best Place to Work in Fintech; Best & Brightest to Work For Nationally; and Comparably’s Best Company Culture, Best Career Growth, Best Engineering Team, and Best Places to Work in Dallas, among others). We’re building a culture where each Alkamist can perform to their highest potential, and we’re always on the lookout for the best and brightest minds. If you’re ready to experience the power of alchemy - transforming the ordinary into the extraordinary - come join one of the fastest growing SaaS companies in the U.S.

As a remote-first company, most of our positions can be remote in the US, except for key roles, which will be indicated in the Job Title.

Follow us on Glassdoor and LinkedIn !

The Sr Platform Engineer is responsible for locating, making visible, and remediating sources of unreliability in the MANTL platform, including correctness problems that surface under failure conditions. This role works directly in the platform's application codebase (TypeScript) and its container-native deployment environment, combining application-engineering skill with reliability-engineering practice. Much of the work is reactive, investigating and resolving issues as they surface, balanced against a standing roadmap of known reliability risks the team has identified and prioritized ahead of time, independent of feature-delivery timelines. This role partners with, but is organizationally and functionally distinct from, both Cloud Infrastructure Engineering and product application engineering teams, focusing specifically on reliability concerns that span or fall between those domains.



Essential Duties & Responsibilities

Investigate, troubleshoot, and resolve reliability issues within MANTL platform application code, including microservice communication failures and correctness issues that emerge under failure conditions

Identify and address failure modes across the platform's third-party and internal system integrations, developing resilience strategies suited to each integration's specific behavior

Design, configure, and maintain monitoring, dashboards, and alerting (Datadog preferred) to increase visibility into platform health and surface emerging issues before they become incidents

Implement and extend distributed tracing across microservices to accelerate root-cause identification for cross-service failures

Diagnose and remediate application performance issues, including caching strategy, inefficient code paths, and query performance

Harden platform and application components against known failure modes through fault-injection and resilience testing, implementing defensive design patterns to prevent recurrence

Build, maintain, and troubleshoot CI/CD build pipelines (GitHub Actions) supporting deployment of the platform

Deploy and troubleshoot container-native (Kubernetes) workloads as part of diagnosing and resolving platform reliability issues

Maintain and execute against a roadmap of known reliability risks, independent of feature-delivery timelines

Create and maintain documentation and runbooks covering platform reliability issues, root causes, and remediations

Act as an escalation point for complex platform reliability issues, partnering with Cloud Infrastructure Engineering and application engineering teams on issues that cross domain boundaries

Contribute to defining reliability targets (SLOs/SLIs) for key platform services

Recommended Experience & Education

Minimum Years of Experience

4 to 7 years of experience in software engineering, platform engineering, or a hybrid development/reliability engineering role

Education Level

Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent work experience

Knowledge, Skills, & Qualifications

Required

Strong proficiency in TypeScript, with production experience in a Node.js/TypeScript service environment

Hands-on experience deploying and troubleshooting containerized workloads in Kubernetes

Experience building and maintaining CI/CD pipelines, GitHub Actions preferred

Experience configuring monitoring, dashboards, and alerting in an APM/observability tool, Datadog preferred

Experience troubleshooting distributed system and microservice communication failures across a variety of integration points and protocols

Experience with distributed tracing tools and practices

Working familiarity with relational databases, sufficient to spot and tune slow or problematic queries

Experience diagnosing and resolving application performance issues (caching, inefficient code paths, slow queries)

Experience designing for and testing failure modes (fault injection, resilience/chaos-style testing) and implementing defensive patterns such as idempotency, retries, and circuit breakers

Strong analytical and troubleshooting skills; ability to work independently on ambiguous, reactive reliability problems

Ability to communicate technical root cause and remediation plans clearly to both engineering and non-technical stakeholders

Preferred

Experience with message brokers or event-streaming platforms such as Kafka

Experience with OpenTelemetry or comparable distributed tracing frameworks

Experience in a regulated or compliance-driven environment (fintech, banking, or similar)

Familiarity with infrastructure-as-code tooling (Terraform or similar)

Experience partnering with infrastructure/SRE teams on issues that cross application and infrastructure boundaries

The salary range for this position is: $145,000 - $165,000

Cool Things to Know

Not Just Any Company : Alkami has an awesome diverse and inclusive environment. We have a FUN culture and offer great benefits, including remote-first environment, unlimited paid time off, 401(k) with employer match, and more.

Work Authorization : We cannot offer employment sponsorship at this time. Candidates must be eligible to work in the US for full-time employment.

Recruiters : We are not looking for outside recruiting firms to help us in this search. Thank you for understanding.

Pay Transparency: As of January 1, 2023, new states and locales have enacted pay equity laws that require more pay transparency by employers in the following states: California, Colorado (effective January 1, 2021), Connecticut, Maryland, Nevada, New Jersey, New York, Ohio, Rhode Island and Washington.

The Important Stuff

Alkami Technology is an Equal Opportunity Employer and Prohibits Discrimination and Harassment of Any Kind: Alkami is committed to the principle of equal employment opportunity for all employees and to providing employees with a work environment free of discrimination and harassment. All employment decisions at Alkami are based on business needs, job requirements and individual qualifications, without regard to race, color, religion or belief, national, social or ethnic origin, sex (including pregnancy), age, physical, mental or sensory disability, HIV Status, sexual orientation, gender identity and/or expression, marital, civil union or domestic partnership status, past or present military service, family medical history or genetic information, family or parental status, or any other status protected by the laws or regulations in the locations where we operate. Alkami will not tolerate discrimination or harassment based on any of these characteristics. Alkami encourages applicants of all ages.

#LI-REMOTE

Original posting on Alkami's site ↗

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job