Jobs and Careers
AL

Staff Platform Engineer (MANTL)

Alkami Technology
Home Office, United States, United StatesRemotefull_timeVerifiedPosted 6 Aug 2026
💰 $175,000/yr($140,000/yr$175,000/yr)

About the role

Alkami is the digital sales and service platform provider for U.S. banks and credit unions. Our unified Platform integrates onboarding, digital banking, and data and marketing—each solution can stand alone, but together they deliver more—to help institutions onboard, engage, and grow relationships. As the future shifts toward Anticipatory Banking, we help data-informed bankers meet the moment with technology that drives action.


Founded in 2009, we continue to be recognized for our intentional culture and tremendous growth (Best Place to Work in Fintech; Best & Brightest to Work For Nationally; and Comparably’s Best Company Culture, Best Career Growth, Best Engineering Team, and Best Places to Work in Dallas, among others). We’re building a culture where each Alkamist can perform to their highest potential, and we’re always on the lookout for the best and brightest minds. If you’re ready to experience the power of alchemy - transforming the ordinary into the extraordinary - come join one of the fastest growing SaaS companies in the U.S.


As a remote-first company, most of our positions can be remote in the US, except for key roles, which will be indicated in the Job Title.


Follow us on Glassdoor and LinkedIn!

The Staff Platform Engineer leads the effort to locate, make visible, and remediate sources of unreliability in the MANTL platform, including correctness problems that surface under failure conditions. This role works directly in the platform's application codebase (TypeScript) and its container-native deployment environment, combining application-engineering skill with reliability-engineering practice at a level of scope and independence beyond the Sr Platform Engineer. In addition to resolving reactive issues as they surface, this role owns and prioritizes a standing roadmap of known reliability risks, driving that work forward on its own timeline. This role partners with, but is organizationally and functionally distinct from, both Cloud Infrastructure Engineering and product application engineering teams, and is expected to set technical direction and standards for platform reliability work.

Essential Duties & Responsibilities

  • Lead investigation, troubleshooting, and resolution of the most complex reliability issues within MANTL platform application code, including microservice communication failures and correctness issues that emerge under failure conditions

  • Set direction for how the platform identifies and addresses failure modes across its third-party and internal system integrations, establishing resilience strategies suited to each integration's specific behavior

  • Own the design and evolution of monitoring, dashboards, and alerting strategy (Datadog preferred) across the platform, ensuring proactive visibility into emerging risks

  • Establish standards and lead implementation of distributed tracing across microservices to accelerate root-cause identification organization-wide

  • Lead diagnosis and remediation of significant application performance issues, including caching strategy, inefficient code paths, and query performance

  • Set direction for platform and application hardening practices, including fault-injection and resilience testing, and drive adoption of defensive design patterns to prevent recurrence

  • Own and improve CI/CD build pipeline architecture (GitHub Actions) supporting deployment of the platform

  • Guide deployment and troubleshooting of container-native (Kubernetes) workloads as part of resolving complex platform reliability issues

  • Define, prioritize, and drive execution of a roadmap of known reliability risks, independent of feature-delivery timelines

  • Establish and maintain documentation and runbook standards covering platform reliability issues, root causes, and remediations

  • Serve as the senior escalation point for complex platform reliability issues, partnering with Cloud Infrastructure Engineering and application engineering leadership on issues that cross domain boundaries

  • Define reliability targets (SLOs/SLIs) for key platform services and advise leadership on reliability risk and tradeoffs

  • Provide technical mentorship and guidance

Apply for this role

Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.

Apply Now →Generate Application Kit

Free account required — sign up in 30s

Company

Alkami Technology

View company profile →