Ambral Logo

Ambral

Head of Research

Posted 4 Days Ago
Be an Early Applicant
In-Office
San Francisco, CA, USA
250K-400K Annually
Expert/Leader
In-Office
San Francisco, CA, USA
250K-400K Annually
Expert/Leader
Own the research agenda for Ambral Labs’ replayable enterprise environment engine. Design experiments, build environment factories, develop graders and evaluation sets, mine tasks from historical workflows, optimize models and agent policies, advance long-horizon post-training methods, and scale replay and observability systems. The role combines hands-on research and production implementation while establishing research culture, recruiting talent, and working directly with the CTO on enterprise deployments.
The summary above was generated by AI
What we do

Ambral Labs helps enterprises own the intelligence behind their most important workflows.

Every company has years of historical evidence showing how work gets done: the context people had, the decisions they made, the actions they took, and the outcomes that followed. Today, most of that history is inert. It isn’t structured in a way that companies can use to evaluate models and improve agent behavior.

Ambral turns this history into replayable environments and eval sets grounded in real workflows and observed outcomes. We use those environments to improve model performance through reinforcement learning and other post-training techniques, alongside context engineering, harness design, and agent engineering.

The result is better, more cost-efficient AI for each enterprise’s specific work, powered by open-weight models that the company can own and control rather than permanently renting from a model provider.
We're YC S2025 , have raised millions in funding, and are already deployed inside multi-billion dollar enterprises. Now we're growing the founding team.

What you’ll do

We’re building a replayable environment engine over real enterprise history.

The system reconstructs a company’s context as it existed at any past time, then exposes that state through the same tools an agent would use in production. This lets us place new policies and agent configurations inside real historical environments, observe how they reason and act, and grade their performance against real outcomes.

As Head of Research, you'll own the research agenda required to make that possible. You'll identify the highest-leverage technical questions, design the experiments needed to answer them, and remain deeply hands-on in building the systems that turn those answers into production.

Some of the problems you'll work on:

  • Building an environment factory that converts recorded enterprise data and task definitions into runnable environments

  • Designing graders that turn ambiguous business objectives into verifiable rewards

  • Developing methods for mining useful tasks, trajectories, and evaluation cases from historical workflows

  • Creating eval sets that are representative, reproducible, and resistant to overfitting

  • Finding the right combinations of models, tools, context, and policies to maximize performance while reducing inference cost

  • Advancing post-training methods for agents that operate over long horizons, incomplete information, and large tool spaces

  • Building replay and observability systems that make agent behavior explainable and measurable

  • Scaling from individual environments to thousands of concurrent training and evaluation runs

These problems are wide open. You’ll have significant ownership over both the research direction and the production systems that make it real.

You’ll work directly with the CTO, deploy into real enterprise workflows, and see your research tested against consequential problems and observable outcomes.

You'll also help establish the research culture at Ambral Labs: how we run experiments, evaluate progress, choose technical bets, and recruit and develop an exceptional research team.

Who you are

You'll likely thrive here if:

  • You have a PhD in machine learning, computer science, mathematics, or an equivalent track record of significant research experience

  • You have deep experience in reinforcement learning, LLM post-training, evals, agent environments, or closely related areas

  • You've taken ambitious, open-ended research problems from hypothesis through experimentation into working systems

  • You’re comfortable turning fuzzy business objectives into tasks and signals that can be evaluated reliably

  • You can move between research questions and production implementation without treating them as separate jobs

  • You’re looking to do the best work of your life and build something you’ll be proud of for decades

We’re especially interested in candidates who have worked at a leading foundation model lab, top AI research organization, or high-performing AI startup.

Benefits
  • Significant equity and ownership

  • Equinox membership

  • Free meals, coffee, and snacks

  • Health insurance

  • Unlimited PTO

Similar Jobs

15 Days Ago
In-Office
Expert/Leader
Expert/Leader
Artificial Intelligence • Enterprise Web • Healthtech • Software
The Head of Research will build and lead Amigo’s clinical AI research function, establish evaluation standards, oversee benchmarks and patient simulations, and design studies demonstrating safe, clinically valid agent behavior. The role includes hiring researchers, managing research programs, publishing findings, developing academic collaborations, engaging clinicians and regulators, communicating system limitations, and translating research results into product and deployment decisions.
Top Skills: Ai AgentsClinical AiClinical Evaluation BenchmarksEhrLlmsPatient Simulation
13 Days Ago
In-Office
200K-250K Annually
Senior level
200K-250K Annually
Senior level
Fintech
Lead quantitative research for a systematic merger arbitrage strategy: develop and backtest signals, maintain the deal database and data quality, apply ML/AI techniques, support portfolio construction and risk decisions, and represent the strategy to institutional clients.
Top Skills: Ai ToolsBacktesting EngineMachine LearningPythonRelational DatabasesSQL
28 Days Ago
Hybrid
249K-289K Annually
Expert/Leader
249K-289K Annually
Expert/Leader
Artificial Intelligence • Fintech • Machine Learning • Mobile • Payments • Retail • Software
Lead a small embedded design and research team to raise polish and consistency across consumer and payments experiences. Build design systems, embed design in discovery with Product/Engineering/Marketing, deepen research practice, apply AI-native scaling, and drive measurable outcomes for onboarding, redemption, and retention.

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account