Mirantis Logo

Mirantis

Senior Storage Systems Engineer - remote in the US

Posted 8 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in USA
Senior level
Remote
Hiring Remotely in USA
Senior level
Deploys, integrates, operates, and tunes high-performance NFS storage for Kubernetes, GPU, and AI workloads. Responsibilities include CSI integration, Linux and network optimization, storage lifecycle management, air-gapped platform support, infrastructure-as-code, GitOps automation, observability, and troubleshooting distributed storage performance and reliability across hybrid, edge, and bare-metal environments.
The summary above was generated by AI
Company Description

About Mirantis

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Job Description

Overview

The role is to deploy, integrate, and operate high-performance storage for GPU-accelerated compute and AI platforms. You will own the storage layer where Kubernetes meets bare metal — standing up NFS-based high-performance storage, wiring it into clusters via CSI, and tuning it to keep data flowing to GPU workloads at scale. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack.

About the Role

We are looking for a senior systems engineer who treats storage as infrastructure to be automated, observed, and tuned — not hand-managed. The right candidate is fluent in Kubernetes storage, deeply versed in Linux storage and networking fundamentals down to the kernel and NFS-client layer, and knows how to make high-performance NAS actually perform under demanding workloads. You should reach for infrastructure-as-code and GitOps by default, be self-directed in diagnosing performance and reliability issues end to end, set operational standards for others to follow, and communicate clearly across teams. Bare-metal hardware experience is a strong plus, but deep Linux storage knowledge is essential.

Responsibilities

1. Storage Integration & Operation

  • Integrate NFS-based high-performance storage (e.g., VAST, Dell PowerScale) into Kubernetes clusters via CSI, storage classes, and persistent volumes.
  • Tune the NFS data path — mount options, nconnect/RDMA, Linux client, and network settings — for high-throughput, low-latency GPU/AI workloads.
  • Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle.

2. Linux Platform & System Integration

  • Configure and optimize Linux systems for storage workloads, including driver setup, file system layout, network tuning, and kernel parameter optimization.
  • Deliver storage integration for k0s-based Kubernetes via Cluster API (CAPI) and K0rdent management/child cluster topologies.
  • Operate storage in fully disconnected (air-gapped) environments, including local artifact/mirror connectivity (Harbor) and PKI/TLS considerations.

3. Automation & Observability

  • Automate storage provisioning and configuration with infrastructure-as-code (Terraform/OpenTofu) and GitOps pipelines (ArgoCD or Flux).
  • Build monitoring, alerting, and observability for storage performance, capacity, and health.
  • Diagnose and resolve performance, reliability, and scaling issues across the storage stack.

Qualifications

Required qualifications:

  • 7+ years of experience in SRE or hardware/storage infrastructure operations
  • 5+ years of building/operating distributed production Storage systems at scale

  • 7+ years of experience in Linux and K8s storage fundamentals (NFS, CSI)

  • 1+ years of experience integrating with or building High Performance Storage solutions (VAST, Weka, DDN, PowerScale)

Additional Information

What does Mirantis offer you?

- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge, open-source innovation;
- Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings, happy hours, hackathons, and tech talks;
- Receive a competitive compensation package with a strong benefits plan.

It is understood that Mirantis, Inc. may use automated decision-making technology (ADMT) for specific employment-related decisions. Opting out of ADMT use is requested for decisions about evaluation and review connected with the specific employment decision for the position applied for. You also have the right to appeal any decisions made by ADMT by sending your request to [email protected]

By submitting your resume, you consent to the processing and storage of your personal data in accordance with applicable data protection laws, for the purposes of considering your application for current and future job opportunities.

We are a Leader for Container Management in G2 (#2 after AWS)!

Similar Jobs

13 Minutes Ago
Remote
TX, USA
Junior
Junior
Insurance • Financial Services
Provide first-tier remote customer support via phone, email, and written correspondence; resolve inquiries promptly, meet quality and performance metrics, handle escalations, and follow operational standards and schedules.
Top Skills: InternetExcelMicrosoft OutlookMicrosoft Word
21 Minutes Ago
Remote or Hybrid
OH, USA
Senior level
Senior level
Financial Services
Lead site reliability engineering for enterprise infrastructure platforms. Build automated, self-healing systems; define and operationalize SLIs, SLOs, observability, and actionable alerting; own production services, incidents, postmortems, security, performance, and cost. Mentor engineers, lead resiliency reviews, reduce toil, and apply validated enterprise AI capabilities to incident response, SDLC workflows, and operational readiness while maintaining security, traceability, and reliability controls.
Top Skills: C++Ci/CdDatadogDynatraceGoGrafanaJavaKubernetesPrometheusPythonRustSplunkTerraform
Senior level
Information Technology • Internet of Things • Mobile • On-Demand • Software
Leads RapidScale’s Risk & Resilience advisory portfolio across enterprise pursuits and client engagements. Responsibilities include executive advisory, solution shaping, proposal and SOW development, pricing and commercial governance, delivery oversight, risk management, practice development, offering creation, seller enablement, and mentoring. The role connects enterprise risk, resilience, cybersecurity, AI, cloud, governance, and compliance into actionable client strategies while ensuring successful transition from sales through delivery.
Top Skills: Artificial IntelligenceAWSCloud ComputingCybersecurityGCPAzureVMware

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account