GoFundMe Logo

GoFundMe

Manager, Machine Learning Engineering

Posted 24 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
219K-329K Annually
Senior level
In-Office
San Francisco, CA, USA
219K-329K Annually
Senior level
Lead and grow an ML/AI operations engineering team responsible for reliable, scalable production systems. Own training pipelines, feature stores, model serving, monitoring, CI/CD, observability, incident response, and generative AI operations. Partner with data science, product, engineering, legal, and privacy stakeholders to establish operational standards, manage vendors, and balance innovation with safety, compliance, cost, and reliability.
The summary above was generated by AI

Want to help us help others? We’re hiring!

GoFundMe is the world’s most powerful community for good, dedicated to helping people help each other. By uniting individuals and nonprofits in one place, GoFundMe makes it easy and safe for people to ask for help and support causes – for themselves and each other. Together, our community has raised more than $40 billion since 2010.

Join GoFundMe as our next Manager, Machine Learning Engineering (ML and AI Operations). In this role, you will lead the team responsible for the infrastructure, pipelines, and operational rigor that keep GoFundMe's machine learning and AI systems reliable, scalable, and safe in production. This role requires strong technical judgment across the ML lifecycle (data → training → online inference → monitoring), a strong understanding of how to enable AI applications to operate safely at scale, and a proven ability to build and lead a high performance team that operates production ML/AI systems with the same rigor as core infrastructure.

Candidates considered for this role will be located in the San Francisco Bay Area. There will be an in-office requirement of 3x a week.

The Job

  • Own the reliability, scalability, and operational health of ML/AI production systems across GoFundMe, including training pipelines, feature stores, model serving, and monitoring/observability infrastructure.
  • Lead, hire, and grow a team of ML/AI operations engineers, setting technical direction through design reviews, architecture decisions, and shared best practices for production ML and AI systems.
  • Partner with data science and ML engineering teams to streamline the path from model development to production deployment, including CI/CD for ML, model packaging, versioning, and rollback strategies.
  • Establish ML operational excellence org-wide by driving standards for model observability (latency, errors, drift, calibration, business KPI deltas), automated retraining triggers, and incident response playbooks.
  • Build and mature on-call processes, SLOs/SLAs, and postmortem practices for ML/AI systems, treating model incidents with the same discipline as production infrastructure incidents.
  • Drive operational strategy for GoFundMe's generative AI systems alongside traditional ML, balancing innovation velocity with safety, compliance, cost, and reliability.
  • Collaborate cross-functionally with Product, Engineering, Design, and Legal/Privacy stakeholders to translate business goals into team priorities and measurable operational outcomes.
  • Manage vendor and platform relationships (e.g., cloud ML platforms, LLM providers) and make build-vs-buy calls that balance cost, control, and speed.
  • Report on team health, system reliability metrics, and operational risk to senior engineering leadership.
  • Employ a diverse set of tools and platforms, including Python, AWS, Databricks, Docker, Kubernetes, Terraform, Snowflake, and GitHub, to guide your team in developing, deploying, and maintaining scalable and robust machine learning systems.

You

  • 7+ years of hands-on experience building and shipping production machine learning systems, with demonstrated ownership of backend services and ML pipelines in a high-availability environment.
  • 1-3+ years of experience directly managing engineers, ideally in an MLOps, ML platform, or infrastructure context, with a track record of hiring and developing strong teams.
  • Strong proficiency in Python and ML libraries/frameworks such as PyTorch, TensorFlow, Scikit-learn, plus strong software engineering fundamentals (testing, code review, CI/CD, API design, performance, and reliability) — enough depth to stay hands-on and credible with your team.
  • Experience designing and operating real-time model serving at scale, including containerization, scalable inference, feature retrieval, and safe rollout strategies (canaries, shadowing, backward-compatible schema evolution).
  • Strong data engineering fluency: building reliable datasets and features using SQL, Spark/Databricks, and warehouse technologies (e.g., Snowflake), with an understanding of event semantics, identity resolution, and data quality controls.
  • Proven experience implementing ML monitoring for both technical and business metrics (drift, calibration, segment performance, latency, error budgets) and running models reliably in production.
  • Familiarity with generative AI/LLM infrastructure and operational considerations (latency, cost, safety guardrails) is a strong plus.
  • Ability to break down ambiguous, high-impact problems, define crisp interfaces and success metrics, and deliver iteratively while managing stakeholder expectations across engineering leadership, product, and data science.
  • Strong leadership and mentoring skills and a proven ability to raise the bar on architecture, engineering quality, and operational rigor for production ML/AI systems.
  • Advanced degree (Master's or Ph.D.) in Computer Science, Statistics, Data Science, or a related technical field is preferred.
  • Sense of humor is optional but appreciated.

Why you’ll love it here

  • Make an Impact: Be part of a mission-driven organization making a positive difference in millions of lives every year.
  • Innovative Environment: Work with a diverse, passionate, and talented team in a fast-paced, forward-thinking atmosphere.
  • Collaborative Team: Join a fun and collaborative team that works hard and celebrates success together.
  • Competitive Benefits: Enjoy competitive pay and comprehensive healthcare benefits.
  • Holistic Support: Enjoy financial assistance for things like hybrid work, family planning, along with generous parental leave, flexible time-off policies, and mental health and wellness resources to support your overall well-being.
  • Growth Opportunities: Participate in learning, development, and recognition programs to help you thrive and grow.
  • Commitment to DEI: Contribute to diversity, equity, and inclusion through ongoing initiatives and employee resource groups.
  • Community Engagement: Make a difference through our volunteering program.

We live by our core values: impatient to be great, find a way, earn trust every day, fueled by purpose. Be a part of something bigger with us!

GoFundMe is proud to be an equal opportunity employer that actively pursues candidates of diverse backgrounds and experiences.  We do not discriminate on the basis of race, color, religion, ethnicity, nationality or national origin, sex, sexual orientation, gender, gender identity or expression, pregnancy status, marital status, age, medical condition, mental or physical disability, or military or veteran status.

The annual U.S. salary range for this full-time position is $219,000 - $329,000. The company also offers equity and other benefits to employees, including healthcare, dental, vision, life insurance and 401(k) saving program. In addition to this wage, there are geolocation differentials that will increase pay depending on the work location. Additionally pay may vary depending on other factors including skills, experience, education, or training. Your recruiter can share more about the specific total compensation package based on your location during the hiring process. 

If you require a reasonable accommodation to complete a job application or a job interview or to otherwise participate in the hiring process, please fill out this accommodation form. 

Global Data Privacy Notice for Job Candidates and Applicants:

Depending on your location, the General Data Protection Regulation (GDPR) or certain US privacy laws may regulate the way we manage the data of job applicants. Our full notice outlining how data will be processed as part of the application procedure for applicable locations is available here. By submitting your application, you are agreeing to our use and processing of your data as required. 

Learn more about GoFundMe:

We’re proud to partner with GoFundMe.org, an independent public charity, to extend the reach and impact of our generous community, while helping drive critical social change. You can learn more about GoFundMe.org’s activities and impact in their FY ‘26 annual report.

Our annual “Year in Help” report reflects our community’s impact in advancing our mission of helping people help each other.

For recent company news and announcements, visit our Newsroom.

HQ

GoFundMe Redwood, California, USA Office

855 Jefferson Ave, Redwood, CA, United States, 94063-9992

Similar Jobs

2 Days Ago
Hybrid
Santa Clara, CA, USA
191K-334K Annually
Senior level
191K-334K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Leads engineering and machine learning teams building enterprise-scale AI Search, retrieval, ranking, recommendation, personalization, and RAG systems. Responsibilities include setting architecture and strategy, deploying production ML models, improving relevance and performance, driving reliability and observability, partnering with product and engineering stakeholders, and mentoring technical teams delivering AI-powered experiences to millions of users.
Top Skills: AlgoliaCloud-Native ArchitecturesCoveoDistributed SystemsElasticsearchEmbeddingsGleanJavaLarge Language ModelsLuceneMicroservicesObservabilityOpensearchPythonRagSemantic SearchSolrVector Databases
25 Days Ago
Remote or Hybrid
San Francisco, CA, USA
240K-340K Annually
Senior level
240K-340K Annually
Senior level
Artificial Intelligence • Big Data • Fintech • Machine Learning
Lead and grow the ML Platform team, owning roadmap, execution, architecture, and production quality for model training, evaluation, serving, and monitoring systems. Partner with customers, Sales, Solutions Engineering, Data Science, Infrastructure, Product, and Customer Success on technical proof-of-value engagements. Modernize feature infrastructure, automate ML lifecycle processes, reduce technical debt, improve evaluation frameworks, and hire and mentor engineering talent.
Top Skills: Apache FlinkAWSDatabricksDockerGCPKafkaKubernetesPythonSpark
One Month Ago
Hybrid
Mountain View, CA, USA
140K-217K Annually
Senior level
140K-217K Annually
Senior level
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Lead a team to measure and improve enterprise search relevance and ranking using ML and LLM/RAG techniques. Manage end-to-end model development, evaluation, deployment, cross-functional collaboration, and team growth to ensure scalable, production-grade search quality.
Top Skills: C++GoLarge Language Models (Llms)PythonRetrieval-Augmented Generation (Rag)

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account