Firecrawl Logo

Firecrawl

Machine Learning Engineer

Posted 20 Days Ago
In-Office
San Francisco, CA, USA
250K-290K Annually
Mid level
In-Office
San Francisco, CA, USA
250K-290K Annually
Mid level
Build, deploy, monitor, and retrain production machine learning models for search ranking, relevance, retrieval, content extraction, classification, and LLM-driven features. Develop large-scale data pipelines from crawl and query data, analyze experiments and behavioral logs, create A/B testing and offline evaluation frameworks, and guide product launch decisions using measurable results.
The summary above was generated by AI
Machine Learning Engineer
 

You'll build the ML behind Firecrawl — the models and the systems that serve them. That starts with search: training and shipping the ranking and relevance models for one of our fastest-growing products, then extending that work across extraction quality and LLM-driven features. You'll also own how we measure: A/B testing launches and building the experimentation frameworks the whole team ships against. If you ship models into production — whether your title says ML engineer or data scientist — this is for you.

 

Salary Range: $250,000–$290,000/year

Equity Range: Competitive equity — details shared during the process.

Location: San Francisco, CA (Onsite)

Job Type: Full-Time

Experience: 3+ years building ML or data-heavy systems in production

Visa: Must be legally authorized to work in the United States. We're not able to sponsor visas right now, though that may change down the line.

About Firecrawl

Firecrawl is the easiest way to turn the web into data AI agents can use. One API call converts any URL into clean, LLM-ready markdown or structured data - the boring-hard problem everyone building with LLMs eventually hits, solved.

We hit 8 figures in ARR in year one and more than doubled it in year two. We have 175k+ GitHub stars, and developers, agents, and category-defining AI companies build on us every day. Growth like this is rare, and we're just getting started.

We're a small team punching far above our weight. Everyone here owns a real piece of the product and company, end to end, and runs it themselves - no hiding behind process or headcount.

This is a place for people who want to work at the frontier: an AI company building the infrastructure other AI companies run on, not one bolting AI onto an existing product. We move fast, go deep, and are building the tools superintelligence will rely on to gather data from the web.

What You'll Do
  • Improve ranking and relevance for Firecrawl Search — from feature engineering to model training to production

  • Build and tune models for learning-to-rank, query understanding, and LLM-driven retrieval

  • Extend ML across Firecrawl's products — extraction quality, content classification, and evaluation of LLM-driven features

  • Mine query logs and behavioral data at scale to find where our products win and where they fail

  • Build the data pipelines that turn web-scale crawl and query data into training data and features

  • Work hands-on with platform, search engineers and cloud DevOps to get models running fast and cheap in production

  • Design and formulate our testing strategy — the A/B testing frameworks and offline evaluation the team ships against

  • Partner on product launches across Firecrawl: define success metrics, run the experiments, and make the ship/no-ship call on evidence

  • Report on how releases perform post-launch and turn the findings into the next iteration

What We're Looking For
  • You've shipped ML models into production systems and owned them after launch — deploying, monitoring, and retraining them, not handing them off

  • You have real ranking or relevance-modeling experience — learning-to-rank, recommendations, or search quality

  • You're comfortable in large, data-heavy systems: query logs, pipelines, and datasets that don't fit in memory

  • You write production-quality code (Python at minimum) and can work inside a real backend codebase

  • You're rigorous about measurement — you've designed and analyzed A/B tests and know when a lift is real

  • You can communicate results clearly to the team — what shipped, what moved, and what to do next

Nice to Have
  • MLOps experience — MLflow, experiment tracking, model registries, or feature stores; Kubernetes is a plus

  • Experience building or standardizing an experimentation framework at a previous company

  • Experience with embedding models, vector retrieval, or LLM-based relevance evaluation

  • Experience evaluating LLM outputs at scale — quality scoring, structured-extraction accuracy, or agent behavior

  • Spark or similar large-scale data processing experience

What We're NOT Looking For
  • A pure statistician or analyst who needs an engineering team to productionize their work

  • Someone who wants to specialize narrowly and hand off everything else

  • Someone who optimizes for process over shipping

A Note On Pace

We operate at an absurd level of urgency because the window for what we're building won't stay open forever. If that excites you, keep reading. If it doesn't, no hard feelings — but this role probably isn't for you.

Benefits & Perks

Available to all employees
  • Salary that makes sense — $250,000–$290,000/year, based on impact, not tenure

  • Own a piece — Gain competitive equity in what you're helping build

  • Generous PTO — 15 days mandatory, anything after 24 days, just ask (holidays excluded); take the time you need to recharge

  • Parental leave — 12 weeks fully paid, for all parents

  • Wellness stipend — $100/month for the gym, therapy, massages, or whatever keeps you human

  • Learning & Development — Expense up to $1,000/year toward anything that helps you grow professionally

  • Team offsites — A change of scenery, minus the trust falls

  • Sabbatical — 3 paid months off after 4 years, do something fun and new

Available to US-based full-time employees
  • Full coverage, no red tape — Medical, dental, and vision (100% for employees, 50% for spouse/kids) — no weird loopholes, just care that works

  • Life & Disability insurance — Employer-paid short-term disability, long-term disability, and life insurance — coverage for life's curveballs

  • Supplemental options — Optional accident, critical illness, hospital indemnity, and voluntary life insurance for extra peace of mind

  • Doctegrity telehealth — Talk to a doctor from your couch

  • 401(k) plan — Retirement might be a ways off, but future-you will thank you

  • Pre-tax benefits — Access to FSAs and commuter benefits (US-only) to help your wallet out a bit

  • Pet insurance — Because fur babies are family too

Available to SF-based employees
  • SF HQ perks — Snacks, drinks, team lunches, intense ping pong, and peak startup energy

  • E-Bike transportation — A loaner electric bike to get you around the city, on us

Interview Process

Application Review — Send us your work and a quick note on why this excites you. Show us what you've built — search systems, indexing pipelines, ranking improvements. We care about what you've shipped, not where you went to school.

Intro Chat (~25 min) — A quick conversation to get to know each other before we go deep. We'll talk about what you've been working on, what drew you to Firecrawl, and what you're looking for in your next role. Time for your questions too.

Technical Chat (~45 min) — We'll dig into a real problem from our world; examples include: improving ranking quality with noisy relevance signals, designing the A/B test for a product launch, or building features from query logs — and talk through how you'd approach it. Come ready to think out loud; we care how you reason, not whether you memorized the answer.

Founder Chat (~25 min) — Culture, pace, ownership, and how you like to work. Time for your questions too.

Paid Work Trial (1-2 weeks) — Work with the team on a real, scoped piece of the product — paid at a contractor rate. It's the truest signal for both sides: you see what building at Firecrawl actually feels like, and we see how you ship. Remote-friendly, and we'll flex around your current commitments.

Decision — We move fast after the trial.

If you want your models ranking results for the whole web — and to see the impact in production the same week — you should join us.

👉 Apply now.

HQ

Firecrawl San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

Yesterday
Easy Apply
Hybrid
San Francisco, CA, USA
Easy Apply
60-60 Annually
Internship
60-60 Annually
Internship
Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Machine Learning Engineer Interns will develop, deploy, and operate production-scale ML models and pipelines. They will lead an end-to-end research-to-production project, apply modern ML techniques to blockchain and crypto use cases, collaborate with senior engineers and product teams, and present findings to stakeholders. The role requires doctoral-level machine learning research, model development experience with PyTorch or TensorFlow, production-quality Python skills, and software engineering fundamentals.
Top Skills: Generative AiPythonPyTorchTensorFlow
5 Days Ago
In-Office or Remote
7 Locations
277K-415K Annually
Expert/Leader
277K-415K Annually
Expert/Leader
Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Build and operate production machine learning systems for ranking, retrieval, recommendations, search, propensity, churn, LTV, and next-best-action decisioning. Design reliable signal contracts covering freshness, provenance, confidence, eligibility, and calibration. Lead feature pipelines, model serving, experimentation, monitoring, and feedback loops while evaluating fairness, risk, compliance, trust, and long-term customer impact. Collaborate across product, growth, data, platform, modeling, risk, and compliance teams.
Top Skills: Ai AgentsBatch PipelinesData LakehousesData WarehousesEmbeddingsEvent StreamsExperimentation SystemsFeature StoresJavaKotlinKubernetesLarge Language ModelsLightgbmModel-Serving InfrastructureObservability ToolingPythonPyTorchRecommendation SystemsSemantic SearchSQLTensorFlowWorkflow OrchestrationXgboost
6 Days Ago
Remote or Hybrid
7 Locations
277K-415K Annually
Expert/Leader
277K-415K Annually
Expert/Leader
Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Build and operate production machine learning systems for ranking, retrieval, recommendations, search, propensity, churn, lifecycle intelligence, and next-best-action decisioning. Design reliable signal contracts with freshness, provenance, confidence, and calibration guarantees. Lead experimentation, monitoring, feedback loops, and impact evaluation focused on fairness, trust, risk, compliance, and long-term engagement. Partner across product, growth, data, platform, modeling, risk, and compliance teams.
Top Skills: JavaKotlinKubernetesLightgbmPythonPyTorchSQLTensorFlowXgboost

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account