Mirantis
Jobs at Mirantis
Let Your Resume Do The Work
Upload your resume to be matched with jobs you're a great fit for.
Success! We'll use this to further personalize your experience.
Recently posted jobs
Software
Design, deploy, and operate high-performance NFS-backed storage for GPU/AI platforms. Integrate storage into Kubernetes via CSI, tune Linux/NFS data paths, automate provisioning with IaC and GitOps, build observability, and diagnose performance and reliability across hybrid, edge, and air-gapped environments.
Software
Design and implement control-plane services, CSI drivers, and tooling to provision and operate high-performance storage for Kubernetes-based GPU/AI workloads. Integrate enterprise and scale-out storage (e.g., PowerScale, VAST), build data-path, telemetry, and observability, and deliver IaC and GitOps pipelines. Support bare-metal, hybrid, and air-gapped deployments via Cluster API and K0rdent, ensuring secure artifact distribution and PKI/TLS.
Software
Design, deploy, and maintain high-performance HPC network fabrics (InfiniBand/RoCE), troubleshoot complex latency/throughput issues, manage Fortinet security solutions, tune and monitor network performance, collaborate with compute/storage teams, document architectures, participate in on-call rotations, and support large-scale GPU/AI clusters and incident triage.
Software
Design and build infrastructure services and APIs that manage bare-metal GPU servers and multi-tenant Kubernetes clusters. Implement provisioning workflows, reconciliation loops, and consoles to visualize inventory and provisioning state while ensuring reliability, idempotency, and durable long-running operation handling.
Software
Own the networking vision and roadmap for k0rdent AI, defining underlay/overlay fabrics, RDMA, DNS/IPAM, network automation, and DPU/SmartNIC integration. Translate customer and engineering requirements into priorities, track emerging interconnect standards, and support positioning, field engagements, and reference architectures.
Software
Design, deploy, manage, and troubleshoot high-performance InfiniBand and Ethernet networks for HPC. Perform performance tuning, capacity planning, and monitoring. Implement Fortinet security, resolve routing/switching/latency issues, collaborate with compute and storage teams, document architectures, and participate in on-call escalation and upgrades.
17 Days AgoSaved
Software
Lead operations for large-scale AI infrastructure: manage NVIDIA GPU and high-performance networking environments, troubleshoot Linux/Kubernetes/storage/hardware issues, drive incident response and root cause analysis, improve observability and automation, mentor engineers, and collaborate with engineering and datacenter teams to ensure platform reliability and scalability.
Software
Develop positioning and messaging for Mirantis AI infrastructure products; translate technical capabilities into audience-specific content; own competitive intelligence; create technical assets (white papers, briefs, reference architectures, sales enablement); support GTM, launches, analyst/media briefings, and industry representation.
Software
Operate and support production AI infrastructure across global datacenters, focusing on high-performance NVIDIA GPU platforms, Kubernetes, and high-speed networking. Monitor and troubleshoot infrastructure, network, hardware, and platform incidents; participate in incident response and root cause analysis; improve observability, automation, and operational runbooks.
