Hybridizer Logo

Hybridizer

DevOps & Infrastructure Engineer (HPC/GPU)

Posted Yesterday
Remote or Hybrid
Hiring Remotely in United States
Entry level
Remote or Hybrid
Hiring Remotely in United States
Entry level
Own and develop HPC/GPU infrastructure for R&D and customer deployments. Design cross-platform CI/CD pipelines, manage on-premise and cloud Kubernetes clusters with GPU passthrough and MIG, configure self-hosted GitHub Actions runners, and maintain DockerHub registries. Assemble and tune physical GPU servers, manage PCIe topology and hardware constraints, and maintain compatibility across NVIDIA drivers, CUDA, and ROCm versions. Provide customer-facing technical assistance for GPU container deployments.
The summary above was generated by AI
Your Mission

You will own the infrastructure that powers our R&D and helps our customers deploy our technology on-premise. You will move beyond standard cloud DevOps into the world of High-Performance Computing (HPC).

  • Think: Design a robust CI/CD strategy that handles cross-platform compilation (Windows/Linux) and execution on specific hardware targets (NVIDIA A100, AMD MI250, Consumer GPUs). Architect solution templates for our customers who need to deploy Hybridizer-generated binaries on their own private clouds.

  • Implement:

    • Set up and maintain Kubernetes clusters (both on-premise and cloud) with GPU Passthrough and Multi-Instance GPU (MIG) configurations.

    • Develop GitHub Actions pipelines that seamlessly dispatch heavy test suites to self-hosted runners equipped with specific GPU accelerators.

    • Configure DockerHub registries and secure container lifecycles for our compiler images.

  • Build:

    • Hardware Tuning: Assemble and fine-tune physical servers. This includes managing PCIe topology, cooling profiles, and power constraints to ensure consistent benchmarking results.

    • Driver Ecosystem: Manage the complex matrix of NVIDIA drivers, CUDA toolkits, and ROCm versions across our fleet, ensuring compatibility with our compiler’s output.

What You Bring to the Table

You are a DevOps engineer who loves hardware. You understand that "the cloud" is just someone else's computer, and sometimes you need to manage that computer yourself.

  • Core DevOps: Strong mastery of Docker and Kubernetes. You know how to write custom Helm charts and manage stateful sets.

  • GPU Infrastructure: You have hands-on experience with NVIDIA Container Toolkit or ROCm integration in containers. You understand concepts like PCIe passthrough, IOMMU groups, and GPU orchestration.

  • CI/CD Automation: Expert in GitHub Actions. You can write complex workflows with matrix strategies and self-hosted runners.

  • System Administration: You are comfortable with Linux kernel tuning, driver installation (dkms), and diagnosing hardware bottlenecks.

  • Customer Facing: You have the communication skills to assist clients. You can explain how to expose a GPU to a Docker container to a sysadmin who might not be an expert in HPC.

  • Adaptability: You are ready to work with a mix of consumer and data-center grade hardware (e.g., configuring a server with 4x RTX 5090s or managing a DGX station).

Similar Jobs

2 Hours Ago
Remote or Hybrid
45K-85K Annually
Junior
45K-85K Annually
Junior
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Handles inbound calls and warm leads to understand customers’ insurance needs, recommend appropriate coverages, and convert prospects into policyholders. The role includes paid training and Property & Casualty licensing, customer communication, sales closing, and brand representation. Employees work remotely, follow assigned evening and weekend schedules, maintain required home-office and internet standards, and remain in their resident state for at least one year.
Top Skills: Cable InternetDsl InternetFiber InternetPc
6 Hours Ago
Remote or Hybrid
83K-139K Annually
Senior level
83K-139K Annually
Senior level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Manages complex commercial energy claims from initial report through resolution, including coverage analysis, liability assessment, reserves, litigation, settlements, and high-dollar exposures. Coordinates defense counsel, experts, legal, underwriting, reinsurance, and claims teams. Advises leadership on coverage and legal developments, mentors claims examiners, oversees litigation strategy, and supports claims process improvement while maintaining regulatory and confidentiality standards.
6 Hours Ago
Remote or Hybrid
7 Locations
185K-327K Annually
Senior level
185K-327K Annually
Senior level
eCommerce • Fintech • Hardware • Payments • Software • Financial Services
Design, build, scale, and operate high-volume payment APIs and distributed systems serving global traffic. Lead complex engineering projects from requirements through production, improve reliability, performance, observability, fault tolerance, and security, and participate in on-call and incident response. Collaborate across product and engineering teams, use AI-assisted development tools responsibly, and mentor engineers through technical reviews and documentation.
Top Skills: AWSGoKafkaKotlinTypescript

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account