Remo Health Logo

Remo Health

Site Reliability Engineer

Sorry, this job was removed at 08:16 a.m. (EST) on Wednesday, Jul 09, 2025
Remote
Hiring Remotely in USA
158K-184K Annually
Remote
Hiring Remotely in USA
158K-184K Annually

Similar Jobs

10 Days Ago
Remote or Hybrid
United States
200K-250K Annually
Senior level
200K-250K Annually
Senior level
Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Lead long-term strategy and architecture for cloud and on‑prem platform infrastructure, driving Kubernetes and multi‑cloud reliability, IaC/GitOps automation, observability, SLO/SLI/error‑budget practices, incident leadership, AI‑augmented tooling adoption, and mentorship of senior engineers to improve platform resilience and developer experience.
Top Skills: Amazon Elastic Kubernetes Service (Eks)AutoscalingAWSCapacity PlanningCi/CdGitopsGoGoogle Cloud PlatformGoogle Kubernetes Engine (Gke)Identity And Access ManagementInfrastructure As CodeKubernetesLinuxNetworkingObservabilityOperatorsPulumiPythonRke2StorageTerraform
13 Days Ago
Remote or Hybrid
110K-145K Annually
Mid level
110K-145K Annually
Mid level
AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Build and maintain automation and reliability for live video distribution across on-prem and cloud. Deploy and manage systems, develop monitoring and automated recovery, troubleshoot complex incidents, coordinate with vendors, document SOPs, support live broadcast components, and participate in L2 on-call rotation.
Top Skills: AacAc3AnsibleAtscAvcAWSBashChefCloudFormationCmafDockerEksGitHevcHlsJavaScriptJSONKubernetesLinuxMicrosoft Graph ApiMpeg Transport StreamsPythonRistScte104Scte224Scte35SrtSsaiSt2022-7St2110StatmuxTerraformUnixXMLYmlZixi
22 Days Ago
In-Office or Remote
Expert/Leader
Expert/Leader
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Define and scale SRE standards across teams, implement SLOs/SLIs/error budgets, build observability and resiliency patterns, drive automation and AIOps, improve reliability for large-scale Azure cloud systems, and influence engineering and platform teams.
Top Skills: Ai/MlAiopsAutomationAzureError BudgetsIncident ManagementLogsObservability (MetricsOpentelemetrySlisSlosTracing)
About Remo

Remo is building a new standard of dementia care by fundamentally changing the care journey for individuals living with dementia and their caregivers (the dyad). As a virtual dementia care provider, our expert clinical team designs personalized, comprehensive care to serve people with dementia and caregiver needs (instead of a one-size-fits-all approach). We empower family caregivers by connecting them with a vibrant community of other caregivers, expert content, and tools to manage the entire dementia journey – from anywhere, at any time. Our mission is simple – to provide accessible, comprehensive, quality dementia care for every person who needs it.

About The Role

We are seeking a highly motivated and experienced Site Reliability Engineer to join our fast-paced healthcare startup. You will be instrumental in ensuring the reliability, performance, and scalability of our modern stack infrastructure. The ideal candidate will have extensive hands-on experience with full-stack development, infrastructure management, and a deep understanding of multi-cloud environments. Experience with FHIR and other medical technologies is highly preferred.

What You’ll Be Doing
  • Ensure the reliability, uptime, and performance of our production systems across multiple cloud environments.

  • Lead and manage incident response, post-mortem analysis, and continuous improvement initiatives.

  • Design, implement, and maintain robust monitoring and alerting systems using Datadog.

  • Develop and automate infrastructure and deployment processes using Terraform and CI/CD pipelines.

  • Collaborate closely with development teams to optimize code quality, deployment efficiency, and system architecture.

  • Help improve our top of the line AI tooling in GCP.

  • Provide hands-on support and troubleshooting for complex production issues, ensuring minimal downtime.

  • Contribute to the design and implementation of scalable, fault-tolerant, and secure systems.

  • Participate in code reviews, architectural discussions, and knowledge sharing sessions.

  • Manage time effectively and drive projects to completion with excellent project management skills.

  • Stay up-to-date with the latest technologies, best practices, and industry trends.

You May Be a Good Fit If You
  • Have at least 6 years of relevant experience as a SRE or Full Stack Engineer . 

  • Have extensive hands-on experience with full-stack development using A React Framework, Node, and Typescript.

  • Have a strong proficiency in Graphql.

  • Are an expert in multi-cloud environments (AWS, GCP) and infrastructure management.

  • Have proven experience with incident response, root cause analysis, and problem resolution under pressure.

  • Have extensive Datadog experience with pipeline, dashboard, and alert experience.

  • Have hands-on experience with CI/CD pipelines and automation tools.

  • Are proficient in Infrastructure as Code (IaC) using Terraform.

  • Are an excellent team player, communicator and have strong problem-solving skills.

  • Have strong time management and project management abilities.

You’re The Ideal Candidate If You Have:
  • 6+ years of experience in Full Stack Engineering.

  • Experience with FHIR and other medical data standards.

  • NextJS Experience.

  • Python with some Data Engineering experience.

  • Deep understanding of serverless architectures and AWS Lambdas.

  • Hasura Cloud and Hasura Enterprise Experience.

  • Prior experience working in a healthcare startup environment.

  • Knowledge of containerization and orchestration technologies (Kubernetes, Docker).

  • Database administration experience (PostgreSQL, MongoDB).

  • Familiarity with security best practices and compliance requirements.

  • Github CI/CD Pipeline optimization skills for Docker, NodeJS, and Vercel. 

  • Expertise in observability tools, particularly Datadog, when it comes to implementation.

Technologies:

  • Kubernetes

  • Graphql

  • Hasura

  • NextJS

  • NodeJS

  • Typescript

  • Python

  • AWS Lambdas

  • Datadog

  • Terraform

  • CI/CD Pipelines

Medical

• 100% Company-paid medical premiums for you and your dependents with HSA options

• Dental and vision plans (50% company-paid premium on employee’s dental plan)

• Dependent care FSA

Financial

• 100% 401(k) match of up to 4%

• $80 / month stipend for cell and wifi

Time Off

• 20 days of PTO and 11 paid holidays

• 5 days sick leave

• 16 weeks fully paid parental leave for birthing parents and 8 weeks for non-birthing parents

• Bereavement leave and pregnancy loss leave

Opt-In Ancillary Options:

• Short-term and long-term disability insurance

• Life insurance

• Critical illness, accident, and hospital indemnity insurance

• Pet insurance

• Legal advice

• Rightway Health, clinical care navigator

• Employee Assistance Program 

Remo aims to reduce health inequities by improving access to affordable, high-quality dementia care. Embracing diversity and equal opportunity are core to that mission--these principles shape our culture, the products we build, and the services we deliver. We celebrate a variety of backgrounds, perspectives, and skills, reflecting the diversity of the caregivers and patients we serve.

We use E-Verify to confirm the identity and employment eligibility of all new hires: Participation Poster (PDF), Right to Work Poster (PDF)

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account