Top Tech Jobs & Startup Jobs in San Francisco Bay Area, CA

5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-250K Annually
Senior level
180K-250K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Build and operate backend services, control planes, and automation for Lightning AI’s managed Kubernetes and Slurm infrastructure. Develop distributed systems for cluster provisioning, lifecycle management, scheduling, observability, reliability, and scalability across large-scale GPU environments. Diagnose complex production issues and collaborate with infrastructure, AI, and platform engineering teams. Contribute to architecture, technical design, mentoring, engineering practices, and on-call operations.
Top Skills: Argo CdCi/CdCloud InfrastructureCrossplaneCustom Resource Definitions (Crds)Distributed SystemsDockerFluxGitopsGoGpu InfrastructureGrpcHelmKubernetesKubernetes ApisKubernetes ControllersKubernetes OperatorsKustomizeNetworkingPythonSlurmTerraform
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
185K-200K Annually
Senior level
185K-200K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Own and scale Lightning AI’s global corporate tax function across a multi-entity organization. Lead federal, state, local, indirect, and property tax compliance; ASC 740 tax accounting; forecasting; audit support; M&A tax; risk management; and process automation. Build policies, controls, documentation, and a scalable tax operating roadmap while partnering with Finance, Accounting, FP&A, Legal, auditors, and external advisors.
Top Skills: CorptaxOnesource
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
100K-150K Annually
Senior level
100K-150K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Support multi-entity general ledger close activities, including journal entries, reconciliations, intercompany, payroll accruals, prepaids, fixed assets, leases, financial statements, and flux analysis. Help transform and document the close function through NetSuite implementation, process redesign, and control development. Support external audits and maintain accurate, auditable accounting records.
Top Skills: ExcelNetSuiteQuickbooks
5 Hours AgoSaved
Remote or Hybrid
San Francisco, CA, USA
170K-210K Annually
Senior level
170K-210K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Designs, deploys, automates, and operates large-scale NVIDIA InfiniBand fabrics for GPU clusters. Manages UFM, switches, firmware, congestion control, routing, QoS, observability, troubleshooting, capacity planning, and incident response. Collaborates with AI platform, GPU infrastructure, storage, and systems teams to build reliable AI Factory environments.
Top Skills: AnsibleBgpEvpnGitGpudirect RdmaGrafanaInfinibandInfrastructure As CodeLinuxMpiNcclNetrisNvidia QuantumNvidia UfmPrometheusPythonQuantum-2Rest ApisRocev2TerraformVxlan
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-250K Annually
Senior level
180K-250K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Design, build, and maintain scalable backend services and APIs powering Lightning AI Studio. Develop platform capabilities for authentication, resource management, multi-tenancy, and developer workflows. Improve reliability, observability, scalability, and performance across cloud-native infrastructure. Collaborate with product, infrastructure, and AI engineering teams, lead technical initiatives from architecture through production, evaluate engineering processes, and mentor engineers.
Top Skills: AWSAzureDockerGCPGoKubernetesPythonReactRust
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-250K Annually
Senior level
180K-250K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Design and build backend systems for AI agent orchestration, tool execution, workflow management, memory, and state handling. Develop scalable APIs and reliable infrastructure for distributed AI workflows, improve platform reliability and observability, and partner with product, research, and engineering teams to deliver agent capabilities. Own technical initiatives from architecture through production, maintain software quality through testing and continuous delivery, and mentor engineers on backend and distributed-systems practices.
Top Skills: APIsAsynchronous ProcessingAutomated TestingAWSAzureCloud-Native ArchitectureContinuous DeliveryDistributed SystemsDockerGCPGoKubernetesObservabilityPythonRust
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-250K Annually
Entry level
180K-250K Annually
Entry level
Artificial Intelligence • Machine Learning • Software
Develop and scale Lightning AI’s backend platform using Go, spanning APIs, infrastructure, billing, security, and integrations. Own features end to end, design reliable and scalable systems, improve architecture and performance, maintain software quality and continuous delivery, reduce technical debt, and mentor engineers. Collaborate with engineering, product, and design teams in a fast-changing SaaS environment.
Top Skills: AWSAzureDockerGCPGoKubernetes
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-250K Annually
Mid level
180K-250K Annually
Mid level
Artificial Intelligence • Machine Learning • Software
Build and deploy production AI systems for customers, translating business objectives into scalable technical solutions. Responsibilities include architecture, proof-of-concepts, software development, deployment, monitoring, debugging, inference optimization, and distributed systems operation. The engineer partners with customer engineering teams, collaborates with product and engineering, improves reusable platform capabilities, and owns technical engagements from discovery through production scaling.
Top Skills: APIsDistributed SystemsDockerGoGpu-Accelerated WorkloadsKubernetesLanggraphModel Serving SystemsPythonRayReactTensorrtTypescriptVector DatabasesVllmWorkflow Orchestration Systems
5 Hours AgoSaved
Remote or Hybrid
San Francisco, CA, USA
180K-220K Annually
Senior level
180K-220K Annually
Senior level
Artificial Intelligence • Machine Learning • Software
Build and operate production software, APIs, tooling, and automation for large-scale GPU, bare-metal, and HPC infrastructure. Responsibilities include provisioning, configuration, monitoring, lifecycle management, observability, hardware integration, reliability improvements, and infrastructure capacity deployment. The role partners with networking, data center, platform, and infrastructure teams to design scalable systems and define technical direction.
Top Skills: APIsBare-Metal InfrastructureBmcContainerizationDell HardwareGpu ServersHpcIpmiJuniper NetworksLinuxOrchestrationPalo Alto FirewallsPxe/IpxePythonRedfishSonicVast
5 Hours AgoSaved
Hybrid
San Francisco, CA, USA
180K-220K Annually
Entry level
180K-220K Annually
Entry level
Artificial Intelligence • Machine Learning • Software
Designs, implements, and maintains secure network infrastructure for Lightning AI’s production and AI environments. Responsibilities include firewall, VPN, IDS/IPS, segmentation, vulnerability assessment, penetration testing, SIEM monitoring, incident response, compliance documentation, access controls, and security automation. The role partners with infrastructure and engineering teams to mitigate threats and optimize security controls.
Top Skills: Cisco AsaCloud SecurityEncryptionEndpoint SecurityFirewallsFortinetIds/IpsNessusNetwork SegmentationPalo AltoQradarQualysSd-WanSecurity AutomationSIEMSplunkVpn
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account