Mirantis Logo

Mirantis

Senior Software Engineer, AI Infrastructure

Reposted 5 Days Ago
Remote
Hiring Remotely in US
Senior level
Remote
Hiring Remotely in US
Senior level
Design and build Kubernetes-based LLM serving infrastructure, including GPU scheduling, deployment, scaling, model lifecycle management, Helm packaging, offline enterprise installs, API integrations, identity, metering, and production observability. Contribute across a multi-service codebase, develop GPU telemetry, write technical design documents, and guide engineering direction on a remote-first senior team.
The summary above was generated by AI
Company Description

Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.  https://www.mirantis.com/

Job Description

Mirantis is building a new enterprise AI infrastructure product that lets organizations run and govern large language models on their own Kubernetes clusters. You will join a small senior team early, with broad ownership of the model-serving layer and its path to production.

What you'll do

  • Design and build LLM serving infrastructure on Kubernetes: deployment, GPU scheduling, scaling, and model lifecycle management.

  • Package the platform for enterprise environments: Helm-based installs, upgrades, and restricted/offline networks.

  • Integrate the serving layer with the platform's API gateway, identity, and metering services.

  • Build the observability for operating GPU inference in production (serving metrics, GPU telemetry).

  • Contribute across a multi-service codebase and help set engineering direction through design docs and reviews.

Qualifications

What we're looking for:

  • 5+ years of software engineering experience in infrastructure, platform, or distributed systems.

  • Deep hands-on Kubernetes experience: building and operating production workloads and Helm charts, not just consuming managed clusters.

  • Experience with GPU workloads or LLM inference, or strong adjacent systems experience and a track record of learning fast.

  • Strong Go programming skills; solid CI/CD and infrastructure-as-code skills.

  • Fluency with AI-assisted development tools (Claude Code, OpenAI Codex) as part of your daily engineering workflow.

  • Comfortable with high autonomy on a small, remote-first, written-culture team.

Nice to have

  • Inference performance work (quantization, batching, caching) or distributed serving frameworks.

  • Enterprise deployment experience: air-gapped installs, SSO/OIDC, supply-chain security.

  • UI development experience (e.g. React/TypeScript), useful as the product's management surfaces grow.

  • Open-source contributions in the Kubernetes or ML-infrastructure ecosystems.

 

Additional Information

What does Mirantis offer you?

  • Work with an established Silicon Valley leader in the cloud infrastructure industry;
  • Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
  • Be a part of cutting-edge, open-source innovation;
  • Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
  • Professional development and training;
  • Attend conferences and working groups;
  • Company outings, happy hours, hackathons, and tech talks;
  • Receive a competitive compensation package with a strong benefits plan.

We are a Leader for Container Management in G2 (#2 after AWS)!

Similar Jobs

4 Days Ago
Remote or Hybrid
CA, USA
160K-322K Annually
Senior level
160K-322K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Deploys and validates complete AI infrastructure software stacks on multi-node GPU systems, then converts implementations into technical documentation, automation, demos, training, and reference architectures. Tests prerelease software, evaluates interoperability and operational resilience, supports partners and field teams, collaborates with engineering and open-source communities, and recommends product improvements based on customer feedback. The role presents solutions through briefings, workshops, webinars, events, and internal training, with some travel required.
Top Skills: Ai InferenceAi TrainingAPIsBare-Metal ProvisioningBluefield DpuCertificate ManagementCi/CdCloud-NativeConfiguration ManagementContainersDgx CloudDocaEthernetGitopsGpu SystemsHelmHpcIdentity ManagementInfinibandInfrastructure As CodeKubernetesLinuxMulti-TenancyNvidia Ai EnterpriseNvidia DgxNvidia DsxObservabilityPythonSecrets ManagementShell ScriptingSlurmStorageTelemetry
16 Days Ago
Remote
USA
Senior level
Senior level
Artificial Intelligence • Software • Automation
Own and operate the realtime voice stack and call-data platform end-to-end: telephony/WebRTC integration, streaming STT/TTS, realtime voice agents, call-to-transcript pipelines, data warehousing, analytics, and dashboards. Build AI-native workflows, RAG/search, human-in-the-loop flows, and optimize performance, latency, and cost. Mentor teammates and collaborate across product and subject-matter experts.
Top Skills: AWSBlandCi/CdClaude CodeConversation IntelligenceCursorData WarehouseDeepgramEvent PipelinesLindyLivekitLlmsN8NPipecatReactReact NativeRetellSipSttTerraformTtsTwilioTypescriptVapiWebrtc
11 Days Ago
In-Office or Remote
184K-357K Annually
Senior level
184K-357K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Design, build, and maintain large-scale AI/ML platform and infrastructure for training, inference, fine-tuning, and Agentic AI. Develop tools, APIs, reliability metrics, and root-cause analyses across application to hardware layers while improving efficiency, resiliency, monitoring, and observability.
Top Skills: C/C++Cloud-NativeDgx CloudDynamoElkGoInfinibandJaxKubernetesLokiNcclNvidia GpusObservability PlatformsPrometheusPythonPyTorchRayRdmaScripting LanguagesTensorFlow

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account