Top Remote Senior Level DevOps & Platform Engineering Jobs in Boston, MA

YesterdaySaved
In-Office or Remote
Boston, MA, USA
126K-162K Annually
Senior level
126K-162K Annually
Senior level
Cloud • eCommerce • Security • Software • Cybersecurity
Design and operate scalable cloud, Kubernetes, networking, CI/CD, automation, observability, security, and reliability solutions. Lead complex troubleshooting, incident response, root-cause analysis, disaster recovery, data platform integration, and infrastructure improvements. Mentor engineers, establish technical standards, reduce technical debt, and partner cross-functionally to improve developer productivity, platform resilience, and release stability. Participate in rotating on-call coverage for production systems.
Top Skills: Ci/CdCloud InfrastructureContainer SecurityData PipelinesDisaster RecoveryDora MetricsGoldengateIamInfrastructure As CodeJenkinsKubernetesLinuxNetworkingObservabilityOciOkeSoc 2
Reposted 16 Days AgoSaved
Easy Apply
Remote or Hybrid
Boston, MA, USA
Easy Apply
119K-170K Annually
Senior level
119K-170K Annually
Senior level
Cloud • Information Technology • Security • Software • Cybersecurity
As a Staff Site Reliability Engineer, you'll oversee Zscaler production data center services, optimize code, and ensure cloud service availability and performance. Collaborate with cross-functional teams to improve processes and resolve escalated issues.
Top Skills: BashDnsFirewallsGrafanaHTTPIcmpLoad BalancingNagiosOsi ModelPrometheusPythonTcp/Ip
One Month AgoSaved
Remote or Hybrid
Boston, MA, USA
140K-215K Annually
Senior level
140K-215K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Manage an engineering team building reusable infrastructure platforms and core distributed systems. Architect highly available services, review designs, plan roadmaps, optimize cloud costs, and enforce deployment standards across AWS and GCP. Lead planning, retrospectives, on-call practices, and reliability initiatives involving SRE, CI/CD, chaos engineering, and incident analysis. Develop engineering strategy, mentor senior engineers, and influence platform decisions across product teams.
Top Skills: AIAnsibleAWSChaos EngineeringChefCi/CdDevOpsGCPGitopsGoLinuxPulumiPuppetPythonSaltSreTddTemporalTerraform
Reposted One Month AgoSaved
Easy Apply
Remote or Hybrid
2 Locations
Easy Apply
127K-249K Annually
Senior level
127K-249K Annually
Senior level
Big Data • Cloud • Software • Database
As a Senior Site Reliability Engineer, you'll design and build complex systems, support Atlas platform operations, automate processes, and ensure high availability of services.
Top Skills: AWSAzureDnsGCPGoHTTPLinuxPythonRubyTls
2 Days AgoSaved
In-Office or Remote
United States
85K-193K Annually
Senior level
85K-193K Annually
Senior level
Automotive
Design, build, and operate a global observability platform across hybrid cloud and on-premises environments. Responsibilities include developing monitoring pipelines, infrastructure-as-code, reliability automation, SLI/SLO frameworks, performance optimization, incident response, root-cause analysis, and production troubleshooting. The role partners with engineering teams to improve system resilience, reduce toil, integrate AI/ML for anomaly detection, and establish observability best practices. It also provides technical mentorship and guidance.
Top Skills: Ai/MlAmazon Web ServicesCC++Ci/CdDatadogDockerDynatraceElkGoGoogle Cloud PlatformJ2EeJavaKafkaKubernetesMicroservicesAzureNagiosNew RelicNoSQLOpentofuPrometheusPythonRestful ApisScalaSensuSplunkSpring BootSQLTcp/IpTerraform
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
3 Days AgoSaved
Remote
US
Senior level
Senior level
Cloud • Software
Own and evolve AWS infrastructure powering a multi-tenant SaaS platform, including EKS, VPC, IAM, RDS, Terraform, Helm, and CI/CD. Lead infrastructure projects, improve reliability and developer self-service, manage SLOs, reduce toil, support SOC 2 compliance, mentor engineers, and participate in incident response. Collaborate with Product teams on reliable customer launches and use observability tools to monitor platform performance.
Top Skills: Amazon EksAmazon RdsAWSCi/CdClaudeCursorGithub ActionsGithub CopilotGoGrafanaHelmIamKubernetesLokiOpentelemetryPrometheusPythonShell ScriptingSoc 2TerraformVpc
6 Days AgoSaved
Remote or Hybrid
17 Locations
160K-210K Annually
Senior level
160K-210K Annually
Senior level
Information Technology • Productivity • Software • Infrastructure as a Service (IaaS)
Own the reliability, performance, architecture, and lifecycle of heterogeneous database platforms. Design and operate relational, NoSQL, caching, time-series, and warehouse systems; improve provisioning, developer tooling, CI/CD, backups, disaster recovery, security, and performance. Influence architecture and product decisions, lead migrations and upgrades, troubleshoot developer and production environments, and establish platform standards, documentation, and regulatory traceability.
Top Skills: Amazon AuroraAmazon RdsAWSAzureBigQueryC++Ci/CdCloudFormationFargateGCPGoIamJavaKotlinKubernetesPostgresRedshiftSnowflakeTerraform
6 Days AgoSaved
Remote
USA
183K-279K Annually
Senior level
183K-279K Annually
Senior level
Consumer Web • Healthtech • Professional Services • Social Impact • Software
Own Headway’s cloud platform and infrastructure engineering strategy. Lead deployment isolation, container orchestration, AWS networking and compute, autoscaling, capacity planning, Terraform self-service infrastructure, cloud cost attribution, and Python runtime performance. Build reliable paved-road tooling that enables engineering teams to deploy safely and independently while improving scalability, observability, and operational resilience.
Top Skills: AutoscalingAWSDatadogDockerEcsEksFastapiIamKubernetesLlm InfrastructureNetworkingPythonRdsTerraform
7 Days AgoSaved
Remote or Hybrid
USA
Senior level
Senior level
Software
Lead DevOps and SRE initiatives focused on Azure SQL Elastic Pools, cloud infrastructure reliability, observability, automation, high availability, disaster recovery, and incident response. Manage scalable Azure or AWS environments, build infrastructure as code and CI/CD pipelines, optimize performance and costs, develop monitoring and remediation solutions, and mentor junior engineers while partnering across development, security, and database teams.
Top Skills: Arm TemplatesAWSAzureAzure CliAzure DevopsAzure Sql Elastic PoolsBicepCi/CdGithub ActionsGrafanaInfrastructure As CodePowershellTerraform
Reposted 7 Days AgoSaved
Remote or Hybrid
United States
101K-170K Annually
Senior level
101K-170K Annually
Senior level
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
Design, operate, and scale SailPoint’s AWS-based Kubernetes infrastructure and service mesh for a global, mission-critical SaaS platform. Responsibilities include cluster architecture, networking, security, GitOps, CI/CD, observability, capacity planning, incident response, compliance, and platform standards. The role influences engineering teams globally, participates in 24x7 on-call support, leads service mesh strategy, and improves reliability, performance, deployment efficiency, and operational practices.
Top Skills: Amazon EksAWSCi/CdFedrampGitopsGoGrafanaHelmInfrastructure As CodeIstioKubernetesLinkerdLinuxNetwork PoliciesOpensearchPci DssPod Security StandardsPrometheusPythonRbacService MeshShell ScriptingTerraform
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account