Maximum of 25 job preferences reached.
Top Remote Senior Level DevOps & Platform Engineering Jobs in Boston, MA
Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Big Data Analytics • Automation
Owns the ServiceNow platform architecture, engineering standards, governance, integration patterns, and technical roadmap. Reviews and approves designs, directs implementation partners, and leads complex solutions across ITSM, CMDB/CSDM, HRSD, Employee Center Pro, asset management, and IntegrationHub. Oversees platform upgrades, AI capabilities, Discovery, and Dynatrace integrations while translating business needs into scalable solutions. Mentors the platform team and guides long-term enterprise service management strategy.
Top Skills:
BoomiCmdbConfluenceCsdmDynatraceEmployee Center ProGenerative AiHamHrsdIntegrationhubIntuneItil V4ItsmJIRAMid ServerNow AssistRest ApisSamSap SuccessfactorsServicenowSharepointSoap ApisVirtual Agent
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Manage an engineering team building reusable infrastructure platforms and core distributed systems. Architect highly available services, review designs, plan roadmaps, optimize cloud costs, and enforce deployment standards across AWS and GCP. Lead planning, retrospectives, on-call practices, and reliability initiatives involving SRE, CI/CD, chaos engineering, and incident analysis. Develop engineering strategy, mentor senior engineers, and influence platform decisions across product teams.
Top Skills:
AIAnsibleAWSChaos EngineeringChefCi/CdDevOpsGCPGitopsGoLinuxPulumiPuppetPythonSaltSreTddTemporalTerraform
Reposted 26 Days AgoSaved
Easy Apply
Easy Apply
Big Data • Cloud • Software • Database
As a Senior Site Reliability Engineer, you'll design and build complex systems, support Atlas platform operations, automate processes, and ensure high availability of services.
Top Skills:
AWSAzureDnsGCPGoHTTPLinuxPythonRubyTls
Fintech • Software
The Senior Site Reliability Engineer ensures SaaS platforms remain reliable, performant, secure, and scalable. Responsibilities include building cloud infrastructure, implementing monitoring and alerting, automating operational runbooks and deployments, managing Infrastructure as Code, applying AI-powered observability and remediation, supporting Kubernetes and cloud networking, and leading incident triage and root-cause analysis during 24/7 on-call rotations.
Top Skills:
AIAiopsAksAnsibleAppdynamicsAWSAzureAzure DevopsBashC# .NetCi/CdCloud NetworkingCloudopsCosmos DbDatadogDynatraceEksFirewallsHarnessIdera Sql Diagnostic ManagerInfrastructure As CodeJavaJenkinsKubernetesLinuxLoad BalancingNew RelicPowershellPythonRedgate Sql MonitorSolarwinds Database Performance AnalyzerSQLTerraformWindows
Healthtech • Information Technology • Software • Telehealth
Develop, monitor, and maintain distributed production systems and AWS-based microservices infrastructure. Build automation, tooling, and repeatable processes that improve uptime, scalability, security, and operational efficiency. Support product engineering teams with performance, scaling, incident diagnosis, and production debugging. Analyze and tune systems, code, and networking while participating in on-call operations and blameless post-mortems.
Top Skills:
AWSDnsDockerGCPGenaiHttp/HttpsKubernetesLoad BalancersNtpReverse ProxiesTcp/IpTlsWeb Application Firewalls
New
Track Smarter, Apply Better.
Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.
Use For Free
Fintech • Legal Tech • Software • Financial Services • Cybersecurity • Data Privacy
Leads the architecture, design, deployment, and operation of enterprise middleware messaging and streaming integrations. Designs multi-region, multi-datacenter platforms with geo-replication, failover, high availability, fault tolerance, and low-latency performance. Evaluates messaging technologies, establishes deployment and monitoring standards, automates infrastructure, collaborates with stakeholders, and mentors junior engineers. The role requires extensive experience with Kafka, RabbitMQ, cloud messaging services, Kubernetes, Docker, clustered environments, and resilient high-volume systems.
Top Skills:
ActivemqApache KafkaAws SnsAws SqsAzure Service BusDockerElasticGcp Pub/SubIbm MqKafka StreamsKubernetesRabbitMQRedpandaStreamnative
Information Technology
Design, implement, administer, and maintain Splunk environments across on-premises and cloud infrastructures. Responsibilities include upgrades, patching, system health monitoring, troubleshooting, data onboarding, log parsing, indexing, normalization, SPL query development, and creation of dashboards, reports, and alerts. The role partners with application and network teams to identify data sources and supports operational, security, and business analytics.
Top Skills:
Regular ExpressionsSearch Processing Language (Spl)Splunk
Security • Cybersecurity
Design, deploy, and maintain Google Cloud infrastructure supporting internal IT. Build repeatable Terraform-based deployments, manage GCP architecture and networking, implement security and compliance controls, monitor performance and costs, troubleshoot cloud environments, and maintain documentation. Collaborate with IT, security, and engineering teams while supporting Azure, AWS, on-premises connectivity, Kubernetes workloads, and infrastructure integration during acquisitions.
Top Skills:
AWSBashClaude CodeCloud InterconnectCloud MonitoringCloud StorageCompute EngineDatadogDnsEntra IdFirewallsGoogle Cloud Platform (Gcp)Google Kubernetes Engine (Gke)IamInfrastructure As Code (Iac)KubernetesLoad BalancingAzureNistPythonRoutingSecurity Command CenterSoc 2SwitchingTcp/IpTerraformVlansVpcVpc Service ControlsVpn
Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Lead a globally distributed IT engineering team building enterprise integrations, workflow automation, orchestration, and AI-native workflows. Set technical direction, establish roadmaps, manage delivery metrics, and partner with business and service operations leaders to eliminate manual work. Guide engineers on architecture, CI/CD, testing, and development practices while evaluating emerging automation and AI tools. Build integrations across identity, ITSM, SaaS, and enterprise systems, and grow team capabilities across US and India time zones.
Top Skills:
AtlassianCi/CdConfluenceGitGoogle WorkspaceJavaScriptJIRAOktaPythonSeleniumSlackTddWorkatoZendesk
Software • Defense
Own reliability, scalability, security, observability, and incident response for production applications across AWS and on-premises DoD environments. Build monitoring and alerting, define SLIs and SLOs, lead post-incident reviews, automate infrastructure with Terraform and Ansible, operate Kubernetes clusters, embed RMF and STIG controls, reduce operational toil, and support secure air-gapped deployments.
Top Skills:
AlloyAnsibleAWSAws GovcloudBashDatadogElk StackGithub ActionsGitlab Ci/CdGitopsGoGrafanaHyper-VIstioJenkinsKubernetesLinkerdLokiNutanixPrometheusProxmoxPythonRmfSecurity+StigsTerraformVMware
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
All Filters
Total selected ()
No Results
No Results





















