Vertex, Inc. Logo

Vertex, Inc.

Principal AI Engineer

Reposted 21 Days Ago
Remote
Hiring Remotely in USA
160K-208K Annually
Expert/Leader
Remote
Hiring Remotely in USA
160K-208K Annually
Expert/Leader
Lead enterprise model-training strategy for Commercial AI products: fine-tune LLMs (QLoRA/LoRA/PEFT), train traditional ML models, design large-scale data pipelines, enforce data governance and PII handling, build reproducible experiment/training pipelines, optimize GPU/distributed training costs, and mentor teams to operationalize models into production.
The summary above was generated by AI

Job Description:

The Principal Engineer, AI Model Training & Data Strategy owns how Commercial AI (CAI) products train, fine-tune, and evaluate models, and how the data behind those models is sourced, curated, stored, and governed. This is primarily a model-training role with a strong secondary focus on the data management and pipelines that make high-quality training possible. The role defines the enterprise training strategy and the standards for how and where training data from Commercial AI products is stored, versioned, and reused. 

Essential Job Functions and Responsibilities 

  • Define and own the end-to-end model training strategy across CAI products, spanning traditional AI/ML models and large language models 

  • Fine-tune large language models using parameter-efficient techniques (e.g., QLoRA, LoRA, PEFT) and full fine-tuning where warranted 

  • Train, evaluate, and tune traditional AI/ML models (classification, regression, ranking, clustering, and similar) 

  • Work with large volumes of data – design and optimize pipelines for ingestion, cleaning, labeling, and feature engineering 

  • Define standards for how and where training data from Commercial AI products is stored, versioned, and accessed (data lakes/warehouses, feature stores, dataset registries) 

  • Establish data governance, lineage, quality, licensing/consent, and PII-handling practices for training data 

  • Build reproducible training pipelines and experiment tracking (datasets, hyperparameters, checkpoints, and metrics) 

  • Define evaluation methodology and benchmarks for model quality, including offline evaluation and regression testing 

  • Curate and clean training, validation, and test datasets, including synthetic data generation where appropriate 

  • Optimize training cost and compute utilization (GPU efficiency, distributed training, quantization) 

  • Partner with product and platform teams to operationalize and hand off trained and fine-tuned models to production 

  • Mentor engineers and raise model-training and data-quality maturity across teams 

Knowledge, Skills, and Abilities 

  • Strong hands-on experience training and fine-tuning both traditional AI/ML models and LLMs in production 

  • Deep experience with parameter-efficient fine-tuning (QLoRA, LoRA, PEFT), quantization, and the tradeoffs versus full fine-tuning 

  • Proficiency with ML/DL frameworks and libraries (e.g., PyTorch, Hugging Face Transformers/PEFT/TRL, scikit-learn) 

  • Experience building and operating large-scale data pipelines and platforms (e.g., Spark, Ray, dbt, or equivalents) 

  • Strong grasp of data management: dataset storage architecture, versioning, lineage, governance, and PII handling 

  • Experience with experiment tracking and reproducible ML (e.g., MLflow, Weights & Biases) 

  • Understanding of distributed training and GPU/compute optimization 

  • Ability to define strategy and standards while remaining hands-on in code 

  • Strong stakeholder collaboration and problem-solving skills 

Education and Experience 

  • Bachelor’s degree in Computer Science, Engineering, or related discipline; advanced degree in ML, AI, or Data Science preferred 

  • 12 or more years of experience in AI/ML engineering, applied ML, or data engineering, with significant hands-on model training and fine-tuning 

Disclaimer 

The above statements describe the general nature and level of work performed in this role. Other duties may be assigned. 


Vertex Values: Together We Win

We're building a team of people who are passionate about making an impact for our customers and committed to how that impact is achieved. Our values define the behaviors, mindset, and culture that make Vertex a great place to grow and do meaningful work.


Play to Win or We Don't Play — If we choose to do something, we're choosing to do it because we plan to win. That mindset raises our bar on product quality, customer outcomes, and how we show up for one another.


Work As a Team, Putting the Customer At the Core — Our customers are our true north. Whatever your role, ask: how will this help a customer succeed today? We earn trust through outcomes, not promises.


Achieve Excellence With Integrity, Speed, and Agility — The market isn't slowing down. We'll move faster, adapt quickly, and never compromise on doing things the right way — for teammates, customers, and partners.


Innovate Boldly With a Growth Mindset — Progress demands smart risk. We'll try new approaches, learn fast, and keep pushing the boundaries — especially where AI can remove friction and unlock value.


Communicate with Care, Candor and Transparency — Honest, constructive conversations make us better. Let's speak plainly about what's working and what isn't and help each other improve.

Pay Transparency Statement:

US Base Salary Range: $159,600.00 - $207,500.00

Base pay offered to new hires may vary based upon factors including relevant industry and job-related skills and experience, geographic location, and business needs.* The range displayed does not encompass the full potential of the role, which allows for further growth and career progression.

In addition, as a part of our total compensation package, this role may be eligible for the Vertex Bonus Plan (VOB), a role-specific sales commission/bonus, and/or equity grants.

Learn more about Life at Vertex and connect with your recruiter for more details regarding Vertex's compensation and benefit programs.

*In no case will your pay fall below applicable local minimum wage requirements.

Similar Jobs

7 Days Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
186K-265K Annually
Expert/Leader
186K-265K Annually
Expert/Leader
Cloud • Information Technology • Security • Software • Cybersecurity
Design and scale a reusable Customer Success AI platform: build foundational AI infrastructure (orchestration, memory, context, workflow engines), optimize model routing, observability, governance, and security, lead architecture and standards, mentor engineers, and partner cross-functionally to deliver enterprise-grade, multi-product AI capabilities.
Top Skills: Agentic LoopsAPIsCloud-NativeContext EnginesDistributed SystemsEnterprise MemoryGovernanceKnowledge GraphsLlm OrchestrationMcp ServersMemory ServicesMulti-Agent SystemsObservabilityOrchestration ServicesPlanning SystemsRagSecuritySemantic ModelsState Management FrameworksVector SearchWorkflow Execution Frameworks
Yesterday
Easy Apply
Remote or Hybrid
United States
Easy Apply
160K-200K Annually
Senior level
160K-200K Annually
Senior level
Cloud • Healthtech • Professional Services • Software • Pharmaceutical
Hands-on technical leader who designs, builds, and deploys production-grade AI automation and agentic workflows. Responsibilities include rapid prototyping, RAG and document-intelligence systems, API integrations, orchestration, monitoring, reusable AI assets, and mentoring a small engineering team to deliver frequent releases and measurable business impact.
Top Skills: AWSAzureClaude CodeCodexGCPGeminiLangchainLanggraphLlamaindexNotebooklmPythonRagVector Databases
37 Minutes Ago
In-Office or Remote
Expert/Leader
Expert/Leader
Artificial Intelligence • Hardware • Software • Semiconductor
Architect and drive a unified, self-service inference control plane and reliability practices for large-scale multi-datacenter and cloud inference fleets. Build capacity orchestration, rollout safety, observability, SLO-based reliability, incident response, and automation; mentor senior SREs and measure impact through reduced toil, faster deployments, and SLO compliance.
Top Skills: BazelCapacity ManagementChaos EngineeringGpu OrchestrationModel Serving RuntimesObservability PlatformsOrchestration SystemsSchedulersWafer-Scale Engine (Wse)

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account