MaintainX Logo

MaintainX

Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in Canada
Mid level
Remote
Hiring Remotely in Canada
Mid level
Partner with product and platform teams to improve reliability, observability, and developer autonomy. Design for reliability, establish standards, build shared tooling, mentor teams on SRE practices, and lead incident response and service health initiatives to scale platform resilience.
The summary above was generated by AI

MaintainX is the world's leading AI-powered maintenance and asset management platform, serving 14,000+ customers including Duracell, Shell, Cintas, and Brenntag. We raised $150M in Series D funding led by Bessemer Venture Partners and Bain Capital Ventures, bringing our total funding to $254M. We were named to the Forbes 2025 Cloud 100, the definitive ranking of the top 100 private cloud companies in the world. We're growing fast and hiring the talent to match.

MaintainX is the world's leading Asset and Work Intelligence platform for industrial and frontline environments. We are a modern, IoT-enabled, cloud-based tool for reliability, safety, and operations of physical equipment and facilities. MaintainX powers operational excellence for 12,000 businesses, including Duracell, Univar Solutions Inc., Titan America, McDonald's, Brenntag, Cintas, Xylem, and Shell.

We recently completed a $150 million Series D funding round, bringing our total funding to $254 million and valuing the company at $2.5 billion.

We’re looking for a Site Reliability Engineer to help advance MaintainX’s reliability, observability, and developer autonomy as we scale our platform.

In this role, you’ll partner closely with product and platform development teams to improve the stability, resilience, and operational readiness of our services. You’ll work alongside teams to design for reliability from the start, establish clear ownership and standards, and build shared tooling that enables teams to operate their services with confidence.

You’ll also contribute to company-wide initiatives that define how MaintainX approaches reliability software development, including observability standards, incident response practices, and service health metrics, helping the organization adopt proven industry practices at scale.

This role is well-suited for an developer who enjoys working across teams, influencing technical direction through strong development practices, and turning reliability principles into practical, scalable systems.

What You'll Do:

  • Assess service maturity and provide insights to development teams

  • Partner with development teams to implement observability best practices

  • Enable development teams to become autonomous with their service deployment, support, and infrastructure

  • Mentor developers on reliability practices, focusing on making them self-sufficient

  • Act as the bridge, ear and eyes of the Platform Division teams to drive tooling and practice adoption across development teams

About You:

  • Deep understanding of observability practices in a distributed system environment and how it influences system design and team behaviour

  • Practical experience with SRE concepts (SLOs, error budgets, incident management)

  • 3–5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems

  • Proficient in cloud-native platforms and infrastructure-as-code concepts and tools

  • Working knowledge of at least one programming language (TypeScript/Node.js is a plus)

  • Excellent communication and collaboration abilities across technical and non-technical teams

  • Ability to translate complex reliability concepts into actionable guidance

  • You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you

What’s in it For You:

  • Competitive salary and meaningful equity opportunities.

  • Healthcare, dental, and vision coverage.

  • 401(k) / RRSP enrollment program.

  • Take what you need PTO.

  • A Work Culture where:

    • You’ll work alongside folks across the globe that reflect the MaintainX values, Smart Humble Optimist.

    • We believe in meritocracy, where ideas and effort are publicly celebrated.

About Us:

Our mission is to deliver one platform for maintenance, repair & operations teams to keep the physical world running. We believe the greatest asset in any organization is the people. That’s why we built an intuitive, mobile-first solution to help boost productivity and collaboration across teams and locations.

MaintainX is committed to creating a diverse environment. All qualified applicants will receive consideration for employment without regard to race, colour, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.

 

Our mission is to deliver one platform for maintenance, repair & operations teams to keep the physical world running. We believe the greatest asset in any organization is the people. That’s why we built an intuitive, mobile-first solution to help boost productivity and collaboration across teams and locations.

MaintainX is committed to creating a diverse environment. All qualified applicants will receive consideration for employment without regard to race, colour, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.

Similar Jobs

2 Days Ago
In-Office or Remote
Senior level
Senior level
Big Data • Information Technology • Software • Analytics • Energy
Manage and scale Enverus' global AWS infrastructure, automate deployments and CI/CD, ensure high uptime, collaborate with developers to enable zero-downtime releases, participate in on-call rotations, and improve operational practices.
Top Skills: AWSAzureC#Ci/CdCloudFormationGoKubernetesLinuxPythonTerraformWindows
Yesterday
In-Office or Remote
Mid level
Mid level
Software
Support deployment, operations, monitoring, and incident response for a production UI service on Kubernetes. Troubleshoot production issues using Splunk and logs, perform first-level debugging of Web Components, assist CI/CD and cloud-native operations, collaborate with engineering teams, and drive reliability improvements against a client-directed backlog.
Top Skills: KubernetesSplunkWeb Components
Mid level
Blockchain • Financial Services • Cryptocurrency • Web3
Operate and improve a shared telemetry platform for metrics, logs, traces, alerting, dashboards, and profiling. Maintain collection, storage, querying, routing, and pipelines; troubleshoot production telemetry issues; deploy and manage services with IaC and container orchestration; build automation for dashboards and alerts; participate in incident response, on-call, and runbook documentation.
Top Skills: AlertmanagerAWSCi/CdClaude (Ai Tools/Agents)ConsulContainer OrchestrationGrafanaGrafana AlloyKubernetesLogqlLokiNomadOpentelemetryPrometheusPromqlPyroscopeSplunkTempoTerraformTerragruntVaultVectorVictoriametrics

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account