Mirantis Logo

Mirantis

Senior DevOps Engineer (Storage) - remote in the US

Posted 14 Days Ago
Remote
Hiring Remotely in USA
Senior level
Remote
Hiring Remotely in USA
Senior level
Design, deploy, and operate high-performance NFS-based storage for GPU/AI Kubernetes platforms. Integrate storage via CSI, tune Linux/NFS/RDMA data paths, automate provisioning with IaC and GitOps, build observability, and troubleshoot performance and reliability across hybrid, edge, and air-gapped environments.
The summary above was generated by AI
Company Description

About Mirantis

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Job Description

Overview

Deploy, integrate, and operate high-performance storage for GPU-accelerated compute and AI platforms. You will own the storage layer where Kubernetes meets bare metal — standing up NFS-based high-performance storage, wiring it into clusters via CSI, and tuning it to keep data flowing to GPU workloads at scale. Work spans hybrid, edge, and air-gapped deployments built on the Mirantis K0rdent stack.

About the Role

We are looking for a senior DevOps engineer who treats storage as infrastructure to be automated, observed, and tuned — not hand-managed. The right candidate is fluent in Kubernetes storage, deeply versed in Linux storage and networking fundamentals down to the kernel and NFS-client layer, and knows how to make high-performance NAS actually perform under demanding workloads. You should reach for infrastructure-as-code and GitOps by default, be self-directed in diagnosing performance and reliability issues end to end, set operational standards for others to follow, and communicate clearly across teams. Bare-metal hardware experience is a strong plus, but deep Linux storage knowledge is essential.

Responsibilities
1. Storage Integration & Operation

  • Integrate NFS-based high-performance storage (e.g., VAST, Dell PowerScale) into Kubernetes clusters via CSI, storage classes, and persistent volumes.
  • Tune the NFS data path — mount options, nconnect/RDMA, Linux client, and network settings — for high-throughput, low-latency GPU/AI workloads.
  • Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle.

2. Linux Platform & System Integration

  • Configure and optimize Linux systems for storage workloads, including driver setup, file system layout, network tuning, and kernel parameter optimization.
  • Deliver storage integration for k0s-based Kubernetes via Cluster API (CAPI) and K0rdent management/child cluster topologies.
  • Operate storage in fully disconnected (air-gapped) environments, including local artifact/mirror connectivity (Harbor) and PKI/TLS considerations.

3. Automation & Observability

  • Automate storage provisioning and configuration with infrastructure-as-code (Terraform/OpenTofu) and GitOps pipelines (ArgoCD or Flux).
  • Build monitoring, alerting, and observability for storage performance, capacity, and health.
  • Diagnose and resolve performance, reliability, and scaling issues across the storage stack.

Qualifications

Required qualifications:

  • 7+ years of experience in SRE or infrastructure operations

  • 5+ years of building/operating distributed production Storage systems at scale

  • Hands-on with High Performance Storage solutions (VAST, Weka, DDN, PowerScale)

  • Linux and Kubernetes storage fundamentals (NFS, CSI)

 

Preferred:

  • Bare-metal experience: hands-on experience with bare-metal host provisioning, raw disk/hardware layout, and physical server storage configurations.

  • Hands-on experience with VAST and/or Dell PowerScale.

  • Experience with GPUDirect Storage and RDMA/RoCE data paths

  • Experience with the Mirantis K0rdent stack (K0rdent Enterprise, K0rdent AI, k0s, MKE) and Cluster API.

  • Familiarity with other storage backends (Ceph, object/S3) and CSI driver operations.

  • Proven experience in sovereign or high-security air-gapped environments.

 

Additional Information

What does Mirantis offer you?

- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge, open-source innovation;
- Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings, happy hours, hackathons, and tech talks;
- Receive a competitive compensation package with a strong benefits plan.

It is understood that Mirantis, Inc. may use automated decision-making technology (ADMT) for specific employment-related decisions. Opting out of ADMT use is requested for decisions about evaluation and review connected with the specific employment decision for the position applied for. You also have the right to appeal any decisions made by ADMT by sending your request to [email protected]

By submitting your resume, you consent to the processing and storage of your personal data in accordance with applicable data protection laws, for the purposes of considering your application for current and future job opportunities.

We are a Leader for Container Management in G2 (#2 after AWS)!

HQ

Mirantis Campbell, California, USA Office

900 E Hamilton Ave,, Campbell, CA, United States, 95008

Similar Jobs

3 Minutes Ago
In-Office or Remote
United States
105K-250K Annually
Junior
105K-250K Annually
Junior
Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
Build and manage a branch-based wealth management practice: discover client goals, create tailored financial plans, drive new business and referrals, coordinate with specialists for investments, insurance, estate planning, and maintain high-net-worth client relationships.
3 Minutes Ago
In-Office or Remote
United States
105K-250K Annually
Junior
105K-250K Annually
Junior
Digital Media • Fintech • Information Technology • Machine Learning • Financial Services • Cybersecurity • Automation
Build and manage a branch-based wealth management practice: perform client discovery, create tailored financial plans, prospect and grow client relationships via partner referrals, collaborate with internal specialists to deliver investment, insurance, and estate planning services, and meet commission-based performance targets.
20 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
Expert/Leader
Expert/Leader
Artificial Intelligence • Cloud • Information Technology • Machine Learning • Software • Big Data Analytics • Automation
Leads strategic technical engagements with enterprise CTOs, CIOs, and engineering executives. Provides architectural oversight for complex opportunities, supports sales cycles, evangelizes PagerDuty’s Operations Cloud, mentors technical sales teams, influences product roadmaps, and builds DevOps/SRE communities. Advises customers on digital transformation, cloud migration, AIOps, incident response, operational excellence, and ethical AI adoption. Represents PagerDuty at industry events and monitors competitive and technical trends.
Top Skills: Ai/MlAiopsAutomationCloud InfrastructureCloud-Native ArchitecturesCybersecurity Incident ResponseDevOpsFedrampItsm/ItilMicroservicesPagerduty Operations CloudSaaSSite Reliability Engineering (Sre)

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account