Filevine Logo

Filevine

Staff Site Reliability Engineer

Reposted One Month Ago
Remote
Hiring Remotely in United States
235K-275K Annually
Expert/Leader
Remote
Hiring Remotely in United States
235K-275K Annually
Expert/Leader
Senior technical leader for SRE driving observability, platform infrastructure, SLIs/SLOs, incident response, automation, and self-service platform capabilities. Shapes reliability strategy, mentors engineers, and ensures production-scale operational excellence.
The summary above was generated by AI
Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. Grounded in a singular system of truth, Filevine brings together data, documents, workflows, and teams into one unified platform—where modern legal work happens with clarity and consistency.
 
Powered by LOIS, the Legal Operating Intelligence System, Filevine connects context across every matter to transform legal operations from reactive to proactive. LOIS reads, understands, and reasons across your data to surface insight, automate complexity, and give professionals the clarity and confidence to see more, know more, and do more. Fueled by a team of exceptional collaborators and innovators, Filevine’s rapid growth has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country.

Role Summary

    As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team
    and a strategic partner to engineering leadership. You don’t just maintain systems — you shape
    engineering culture, define the technical standard for how Filevine runs in production, and bridge the
    gap between high-level business goals and robust, internet-scale technical execution. You bring a
    forward-looking perspective — actively shaping how AI and machine learning drive the future of
    reliability practice.

    You own the roadmap across two critical SRE domains — Observability & Alerting and Platform
    Infrastructure — and are accountable for ensuring the team solves reliability problems permanently
    rather than absorbing them as toil. You operate as the senior IC counterpart to the Engineering
    Manager: technical correctness lives with you. You partner with the Reliability Architect and engineering
    leadership on significant technical decisions, mentor engineers across experience levels, and influence
    reliability strategy across the broader organization. Reliability at Filevine protects revenue. You are the
    senior technical voice responsible for ensuring that uptime, incident response, and every production
    change meet the operational standard the business demands.

    This role does not participate in on-call rotation, but you are deeply invested in the engineers who do —
    shaping the on-call strategy, tooling, and culture that make production support sustainable and
    effective.

Who You Are

    The Technical Authority
    Master of the Craft: You bring deep expertise in distributed systems, cloud infrastructure,
    observability, and reliability engineering. You raise the technical standard for every engineer
    around you and thrive where the challenges are complex and the stakes are real.
    Technical Leader and Mentor: You are passionate about mentoring engineers and investing in
    their growth. You influence technical direction and communicate production risk clearly across
    engineering, product, and executive audiences.
    Forward-Thinking & AI/ML Fluent: You bring deep knowledge of AIOps and drive the use of
    AI and machine learning in observability, anomaly detection, incident response, automated
    remediation, and resource optimization.
    Production-Scale Problem Solver: You turn ambiguous, complex reliability challenges into
    durable solutions for systems where availability, performance, and production changes carry
    meaningful business impact.
    Software-Minded Builder: You use software, automation, Infrastructure as Code, and platform
    capabilities to eliminate toil and make systems safer, more scalable, and easier to operate.
     
     

What you will do

  • Define and execute the technical strategy for Observability & Alerting, Platform Infrastructure,
    and operational excellence.
  • Lead the evolution of reliable, scalable, secure, and efficient cloud platforms and distributed
    systems.
  • Champion SLIs, SLOs, error budgets, capacity planning, operational readiness, and automation
    across the service lifecycle.
  • Lead the organization through complex production incidents and turn post-incident learning into
    permanent engineering improvements.
  • Build self-service platform capabilities that reduce toil, improve engineering safety and velocity,
    and make every team more capable of owning their own reliability.
  • Mentor engineers and serve as a trusted technical authority for long-term reliability and platform
    direction.

Qualifications

  • 12+ years of experience in software engineering, infrastructure, platform engineering, or SRE,
    including 6+ years in SRE and 3+ years leading complex, cross-functional technical initiatives
    for distributed production systems.
  • Expert-level depth in observability and platform infrastructure, with broad expertise in incident
    response, capacity planning, automation, and reliability engineering.
  • Advanced experience with a major container-orchestration platform, preferably Kubernetes, and
    an observability platform such as New Relic, Datadog, or equivalent.
  • Strong software-engineering ability in Python, Go, Bash, or another general-purpose language,
    with experience building production tooling, automation, or platform capabilities.
  • Proven ability to mentor engineers and communicate technical risk clearly to engineering,
    product, and executive audiences.
  • Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is
    strongly preferred.

Compensation Information: $240,000 - $280,000
 
The base salary range represents the low and high end of the salary range for this position. The total compensation package for this position will be determined by each individual’s location, qualifications, education, work experience, skills and performance. We believe in the importance of pay equity - the range listed is just one component of Filevine’s total compensation package for employees. This position is also eligible for stock options, a paid time off policy, as well as a comprehensive benefits package.
 

Work Location Expectation: Remote or option to be hybrid/in-office in one the following locations: San Francisco, New York, Chicago, or Salt Lake City
 
 
Cool Company Benefits:
- A dynamic, rapidly growing company, focused on helping organizations thrive 
- Medical, Dental, & Vision Insurance (for full-time employees)
- Competitive & Fair Pay
- Maternity & paternity leave (for full-time employees)
- Short & long-term disability
- Opportunity to learn from a dedicated leadership team
- Top-of-the-line company swag
 
Privacy Policy Notice
Filevine will handle your personal information according to what’s outlined in our Privacy Policy.
 
Communication about this opportunity, or any open role at Filevine, will only come from representatives with email addresses using "filevine.com". Other addresses reaching out are not affiliated with Filevine and should not be responded to.
 

Similar Jobs

3 Days Ago
Remote or Hybrid
Senior level
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Staff Software Engineer responsible for operating production Kubernetes across hybrid and multi-cloud environments, building auto-remediation and observability systems, developing Infrastructure-as-Code and GitOps pipelines, defining SLOs and alerting, supporting incident response, and reducing operational toil. The role also includes mentoring SRE engineers, improving reliability practices, and contributing to cloud, data center, and on-call operations.
Top Skills: AksAWSAzureBashCi/CdCloudFormationEc2EksGCPGitopsGkeGoInfrastructure As CodeKubernetesLinuxMachine LearningPythonRdsService MeshTerraform
12 Days Ago
Easy Apply
Remote
United States
Easy Apply
Senior level
Senior level
Cloud • Security • Software • Cybersecurity • Automation
Build and operate reliable, scalable production infrastructure for GitLab’s user-facing services. Responsibilities include developing infrastructure automation and tooling, managing Kubernetes deployments, maintaining infrastructure as code, supporting CI/CD and GitOps, participating in on-call and incident response, improving observability and SLOs, troubleshooting production systems, and documenting operational practices. The role spans Intermediate through Senior Staff levels and requires strong software engineering, cloud, reliability, and asynchronous collaboration skills.
Top Skills: AlertingAWSCi/CdGCPGitopsGoInfrastructure As CodeKubernetesLoggingMetricsRubySlisSlosTerraform
2 Days Ago
Remote
United States
250K-325K Annually
Senior level
250K-325K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Software • Database • App development • Generative AI
Lead reliability engineering for Replit’s large-scale infrastructure by designing observability, defining SLOs and SLIs, leading incident response, automating operations, optimizing Kubernetes and GCP deployments, debugging distributed systems, and mentoring engineers. Build internal tools and integrations in Python or Go, maintain infrastructure as code and CI/CD pipelines, improve system performance and resilience, and establish reliability, security, and operational best practices across the engineering organization.
Top Skills: Ci/CdDatadogDockerGoGoogle Cloud Platform (Gcp)GrafanaKubernetesOpentelemetryPrometheusPulumiPythonTerraform

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account