fal Logo

fal

Software Engineer, Virtualization

Reposted 14 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
180K-250K Annually
Senior level
In-Office
San Francisco, CA, USA
180K-250K Annually
Senior level
Build, provision, and automate high-performance bare-metal and virtual compute environments with GPU passthrough and dedicated Kubernetes clusters. Design overlay networking and routing, maintain Linux images, implement monitoring, and automate lifecycle, leveraging AI for provisioning, alerting, and recovery.
The summary above was generated by AI

fal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.

As generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.


About this role: 

You build the custom compute environments we deliver to customers — bare metal or virtual machines with GPU passthrough, dedicated Kubernetes clusters, and the networking that ties them together. You work across the full stack from Linux image building to overlay network design to cluster bootstrapping.

Key responsibilities
  • Build and deliver custom environments with excellent GPU performance for customer workloads
  • Leverage AI to an extreme level to automate provisioning, alerting and recovery
  • Provision and configure dedicated Kubernetes clusters tailored to customer requirements
  • Design and implement overlay networking (VLAN, VXLAN) and routing configurations (ECMP, BGP) and tunnels (strongSwan, IPSEC) for tenant isolation and performance
  • Build and maintain Linux images
  • Set up network monitoring and diagnostics for customer environments
  • Automate the end-to-end lifecycle of customer compute environments: creation, configuration, validation, and teardown
Requirements
  • 5+ years experience with Linux virtualization: KVM/QEMU, libvirt, VFIO device passthrough, hugepages, NUMA, CPU pinning
  • Strong networking fundamentals: VXLAN, VLAN, ECMP, BGP, ARP, and the ability to debug packet-level issues (tcpdump, Wireshark)
  • Production experience building and operating Kubernetes clusters on bare metal (MetalLB)
  • Proficiency with Linux image building and OS provisioning (kickstart, cloud-init, PXE/iPXE)
  • Proficiency in Python, Bash, Ansible and Terraform
  • Deep experience with NVIDIA GPUs: drivers, MIG, container runtimes (nvidia-container-toolkit), InfiniBand, RDMA/RoCEv2 and GPUDirect for high-performance AI networking
  • Excellent communication and ability to drive technical decisions across teams
  • Self-starter who executes quickly, takes ownership, and constantly seeks improvement
Nice to have
  • Experience with SR-IOV, DPDK, or other high-performance networking technologies
  • Experience with shared network storage (Ceph, Lustre, Weka)
  • Experience with network automation tools (Netbox, Nautobot, Nornir)
Compensation
  • $180,000-250,000 plus equity + benefits (This range encompasses 2 levels Senior and Staff)
Location
  • San Francisco, CA

What we offer at fal
  • Interesting and challenging work
  • A lot of learning and growth opportunities
  • We are currently hiring in downtown San Francisco.
  • We offer relocation assistance to San Francisco.
  • Health, dental, and vision insurance (US)
  • Regular team events and offsites

Similar Jobs

2 Days Ago
Hybrid
Palo Alto, CA, USA
137K-182K Annually
Expert/Leader
137K-182K Annually
Expert/Leader
Automotive • Cloud • Hardware • Software
Develop and maintain a user-space runtime that enables production ECU firmware to run on workstations and cloud servers. Model peripherals, integrate RTOS-based firmware with host abstractions, maintain cross-compilation builds, and extend HIL Pytest validation suites. Debug firmware-host discrepancies using GDB and protocol analysis, collaborate across hardware and software teams, lead design reviews, improve code quality, and mentor engineers on real-time embedded development.
Top Skills: BazelCC++CanCertDoipEthernetGdbHilLinMisraPosixPytestPythonQemuRenodeRtosSocketcanSome/IpSynopsys Virtualizer Development Kit
21 Days Ago
Hybrid
Menlo Park, CA, USA
157K-230K Annually
Mid level
157K-230K Annually
Mid level
Artificial Intelligence • Big Data • Cloud • Machine Learning • Software • Database • Analytics
Build and improve an AI-native virtualization platform that makes applications work on Snowflake. Work spans AI-driven software synthesis, query parsing/optimization, wire protocols, and data ingest/egress. Collaborate with product, solutions, and field engineering to enable customer migrations and define AI-based engineering processes.
Top Skills: AIData EgressData IngestDatabase InternalsFunctional Programming LanguagesHigh-Performance ComputingQuery OptimizationQuery ParsingSnowflakeSQLWire Protocols
17 Days Ago
In-Office
2 Locations
139K-242K Annually
Senior level
139K-242K Annually
Senior level
Cloud • Information Technology • Machine Learning
The Senior Software Engineer will design and build secure sandboxed runtime environments for Kubernetes, focusing on GPU workloads and ensuring performance and security in multi-tenant configurations.
Top Skills: BashC/C++GoGpuGvisorKata ContainersKubernetesKubevirtLinuxQemuRust

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account