Vectara Logo

Vectara

Platform Engineer

Sorry, this job was removed at 11:10 a.m. (PST) on Tuesday, Apr 28, 2026
In-Office or Remote
Hiring Remotely in San Francisco, CA, USA
In-Office or Remote
Hiring Remotely in San Francisco, CA, USA

Similar Jobs

4 Hours Ago
Remote or Hybrid
USA
150K-170K Annually
Mid level
150K-170K Annually
Mid level
eCommerce • Fintech • Food • Mobile • Social Impact
Build, operate, and evolve AWS cloud infrastructure supporting a high-growth financial and hospitality technology platform. Responsibilities include infrastructure architecture, deployment pipelines, reliability engineering, incident response, observability, security, governance, capacity planning, migrations, and on-call support. The role partners with application engineers, implements infrastructure changes, and improves operational standards and scalability.
Top Skills: AlbAlertingAWSCloudFormationCloudwatchContainersEc2EcsEksElasticacheFargateIamInfrastructure As CodeLogsMetricsNlbRuby on RailsRdsRedisTerraformTracingValkeyVpc
2 Days Ago
Easy Apply
Remote
United States of America
Easy Apply
200K-220K Annually
Senior level
200K-220K Annually
Senior level
Information Technology • Cybersecurity
Build, monitor, and scale resilient cloud infrastructure across multi-region and multi-cloud environments. Develop Kubernetes controllers and infrastructure tools using Go, Python, Ruby, and Terraform; maintain Datadog observability; diagnose complex production performance and reliability issues; automate failover and capacity management; and collaborate with engineering teams on architecture and infrastructure improvements.
Top Skills: Ai Coding AgentsCloud InfrastructureDatadogGoKubernetesLinuxMulti-Cloud InfrastructurePythonRubyTerraform
13 Days Ago
Easy Apply
Remote
USA
Easy Apply
145K-170K Annually
Junior
145K-170K Annually
Junior
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Own and operate the production threat intelligence platform, maintaining system health, monitoring jobs, and ensuring data quality. Build intelligence-feed integrations, enrichment, tagging, and lifecycle workflows. Deliver intelligence to detections, blocklists, warehouses, and agent workflows. Partner with Security and Product stakeholders to translate requirements into technical capabilities, troubleshoot production issues, and implement durable fixes. Develop production-quality software using Python, Go, or Storm with CI/CD ownership, while responsibly applying generative AI.
Top Skills: Ci/CdGoPythonStormVertex Synapse

Platform Engineer

Vectara is the Enterprise Agent Platform that enables businesses to build and deploy governed, grounded, auditable AI agents across SaaS, VPC, and on-prem. We have developed our own models, observability and orchestration layers, guardrails, and context management systems in order to provide our customers with the greatest levels of AI agent accuracy, security, and explainability.
Our founding members include industry veterans and experts in neural information retrieval and distributed systems from Google. We’re a passionate team that’s hyper-focused on solving enterprise-level technology and business problems with AI. Join us!

The Role
You'll own the infrastructure that runs our deploy anywhere platform — from Kubernetes clusters serving ML inference at scale to the CI/CD pipelines, IaC, and observability stack that keep it all reliable. This is a hands-on role: you'll write Helm charts and Terraform one day, debug a Kafka consumer lag issue the next, and ship a backend service feature the day after. You'll deploy across AWS, GCP, and on-premises (including air-gapped environments), and you'll participate in an on-call rotation supporting enterprise customers.

What You’ll Do

  • Build and maintain infrastructure-as-code (Terraform, Helm) for our AWS EKS and GCP GKE clusters, plus on-premises deployments (including Tanzu and air-gapped environments).
  • Own CI/CD pipelines (GitHub Actions, Bazel, ArgoCD) and drive GitOps adoption.
  • Deploy, scale, and optimize ML/NLP inference workloads (vLLM, PyTorch, GPU scheduling with various Kubernetes scalers).
  • Build and improve observability: Prometheus, Grafana, Datadog,, and OpenTelemetry.
  • Collaborate with Field Engineering to support PoCs and platform deployments in customer cloud VPCs and on-prem environments.
  • Contribute to backend services (Java 21, Python, gRPC) and platform features.
  • Improve system reliability, scalability, and developer experience across the engineering org.

What You’ll Bring (Required):

  • 2+ years in platform engineering, DevOps, SRE, or backend infrastructure roles.
  • Strong Kubernetes experience (deployment, debugging, scaling — not just `kubectl apply`).
  • Hands-on with infrastructure-as-code: Terraform, Helm, or Pulumi.
  • Experience with at least one major cloud provider (AWS preferred; GCP or Azure also valued).
  • Proficiency in one or more of: Go, Python, Java. Comfortable reading and contributing to backend codebases.
  • Working knowledge of CI/CD systems (GitHub Actions, Bazel, ArgoCD, or similar).
  • Solid fundamentals in Linux, networking, and distributed systems.

What Sets You Apart (Preferred)

  • Experience deploying or operating ML inference workloads (model serving, GPU scheduling, vLLM, TensorFlow Serving, or similar).
  • Familiarity with streaming/messaging systems (Kafka, Pulsar) and data stores (MariaDB/PostgreSQL, Aerospike, ClickHouse, OpenSearch).
  • Experience with GitOps workflows (ArgoCD, Flux).
  • Exposure to air-gapped or on-premises Kubernetes deployments.
  • Background in observability tooling (Prometheus, Grafana, OpenTelemetry, Datadog).
  • Experience providing technical support or working directly with enterprise customers on infrastructure issues.
  • Comfort with AI-assisted development workflows and managing AI coding agents.

Location requirements:

We support remote applicants from all over the US but candidates who can come to the office 2-3 days a week in our Palo Alto office are preferred. 

Perks and Benefits:

100% paid Medical, Dental, Vision for employees.  Option of Health Savings Account (HSA) or Flexible Savings Account (FSA). Generous paid time off (PTO) plus paid sick time and holidays. Professional development and training opportunities. Company virtual happy hours and fun team building activities and more. 

Salary is just one component of Vectara’s employee compensation. Our full-time employees are also equity owners in the company, which although not an immediate cash component, can have positive impacts on long-term total compensation for each participating employee. We would be remiss if we didn’t highlight and celebrate our focus on engaging many of our employees in being economic co-owners of the business.

Vectara welcomes all. We value the collective wisdom of people from different backgrounds, experiences, abilities and perspectives.  We never discriminate on the basis of race, religion, national origin, gender identity or expression, sexual orientation, age, or marital, veteran, or disability status. Vectara has a positive and supportive culture—we look for people who are inventive and work to be a little better every single day. We seek to be smart, humble, hardworking and, above all, curious. After all, we are on a mission to find meaning.

HQ

Vectara Palo Alto, California, USA Office

395 Page Mill Road Ste 275, Palo Alto, CA, United States, 94306

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account