Hilbert's AI Logo

Hilbert's AI

AI Engineer - Core

Reposted 12 Days Ago
Hybrid
San Francisco, CA, USA
140K-170K Annually
Mid level
Hybrid
San Francisco, CA, USA
140K-170K Annually
Mid level
Design, build, and maintain production-grade AI systems and agent-based workflows; own end-to-end pipelines from experimentation to deployment and monitoring; build evaluation pipelines; collaborate with product, data, and GTM; iterate quickly under ambiguity to shape the AI stack.
The summary above was generated by AI

Hilbert is building a reasoning engine that must navigate non-deterministic user behavior across data silos — turning months-long decision cycles into minutes. Fully agentic by design, our demand intelligence platform doesn't just call APIs; it solves the hard problem of orchestrating multi-step inference over messy, high-stakes enterprise data where deterministic answers don't exist.

From Fortune 500 enterprises to beloved brands like FreshDirect, Blank Street, and Levain Bakery, operators run their growth on Hilbert. We're also co-building alongside leading AI companies.

We're looking for an AI Engineer who can build production-grade AI systems end-to-end — from prototype to pipeline to product — with the ownership and urgency of a startup culture.

This is not a "wire up a prompt chain and move on" role. You'll own core pieces of the AI stack that power Hilbert's demand intelligence platform — designing agent architectures, building evaluation systems, and making hard tradeoffs between accuracy, latency, and cost in production. You'll ship fast in conditions where the spec is evolving, and communicate what you're building (and why) with clarity to the rest of the team. If you think in systems, have opinions about how agentic workflows should actually work, and want to build AI products that drive real enterprise outcomes, we want to meet you.

THE ROLE

You'll work directly with the founding team and across product, data, and GTM to design, build, and improve the AI systems at the heart of Hilbert. The environment is high-autonomy and high-ambiguity — the nature of building AI-native products means requirements shift, approaches evolve, and the person closest to the problem often makes the call.

What you'll do:

  • Design, build, and maintain AI-driven features and pipelines that serve enterprise customers at scale

  • Architect and implement agent-based workflows using LangChain, LangGraph, or equivalent orchestration frameworks

  • Own systems end-to-end — from experimentation through production deployment and monitoring

  • Build and improve evaluation pipelines to measure, validate, and iterate on AI system performance

  • Collaborate closely with the founding team and cross-functional partners — communicating tradeoffs, progress, and technical decisions with clarity

  • Make pragmatic engineering decisions under ambiguity — ship, learn, iterate

  • Shape the technical direction of the AI stack as the company scales

Our Current Hurdles

These are the kinds of problems you'll walk into on day one:

  • Intelligent retrieval across heterogeneous approaches — our agents need the right information at exactly the right moment. The challenge isn't picking one retrieval method; it's combining RAG, graph-based retrieval, and other approaches into a unified strategy that fetches the most relevant content precisely when the agent needs it — no more, no less.

  • Agentic workflows that solve real-world problems — it's building workflows robust enough to handle the unexpected. When an agent hits an edge case, missing data, or a situation it wasn't explicitly designed for, it needs to reason through it — leveraging available context, escalating to a human when it can't, and never silently failing.

  • Evaluation beyond vibes — we need systematic, reproducible evals that actually predict real-world performance. If you've built custom evaluators for RAG or agent workflows, we want to talk.

  • Execution and real-world integration — an agent that only surfaces insights isn't enough. We're building systems where agents take action — integrating with external platforms, executing workflows, and doing real work with the information they have, combined with human-in-the-loop checkpoints that keep enterprise trust intact.

WHO THRIVES IN THIS ROLE

We care about how you think and how you ship - not how many years are on your resume.

The profile:

  • You're a strong Software engineer. Your code is clean, testable, and production-ready.

  • You have real experience with LangChain, LangGraph, or equivalent agent/orchestration frameworks. You've built with them, hit their limits, and worked around them - not just followed tutorials

  • You communicate with clarity and conviction. You can explain a technical decision to a non-technical founder and debate architecture tradeoffs with a senior engineer . Communication is not a nice-to-have here - it's core to the role

  • You take ownership. You don't wait for tickets. You see what needs to be built, raise your hand, and ship it

  • You thrive in ambiguity. AI products evolve fast. Requirements change. You're energized by figuring it out.

  • You move at startup speed. You understand what it means to be available, responsive, and biased toward action in a fast-moving, early-stage environment

Strong pluses:

  • Experience building evals pipelines — designing metrics, running systematic evaluations, and using results to drive iteration on AI systems

  • Backend software engineering experience — building APIs, services, data infrastructure, or production systems

  • Exposure to retrieval-augmented generation (RAG), vector databases, or LLM-powered search and recommendation systems

  • Experience at early-stage startups or high-growth environments where you wore multiple hats

You might be:

A backend engineer who went deep on LLMs and never looked back. An ML engineer who realized they love building products, not just models. A startup CTO who wants to go deep on AI at a company where the stack is the product. Someone who's been hacking on agents and pipelines nights and weekends and wants to do it full-time with real enterprise stakes. What matters: you ship, you own it, and you communicate like a teammate — not a silo.

Location

San Francisco, with occasional travel for team meets, offsites or customer engagements.

The Hiring Journey

Short form → Intro call → Technical working session → Team conversations → Offer

Why join us

At Hilbert, we move fast, work collaboratively, and give people real ownership over their impact. As a fast-growing company, there's no shortage of room to grow — you'll take on new challenges quickly and shape the path as you go.

On top of that, here's how we take care of our team:

Health & wellness. We cover 100% of your Health, Dental, and Vision premiums — with the option to add coverage for dependents or a spouse at additional cost.

Financial future. Plan ahead with our 401k through Human Interest, including an employer match that vests immediately.

Time to recharge. Enjoy generous PTO so you can rest, travel, and spend time on what matters most outside of work.

Commuter support. Based in SF? We help cover your commuting costs.

Lunch, covered. In the SF office, we provide lunch vouchers via DoorDash.

And many more. We're always looking for new ways to support our team. Come build the future with us, and grow alongside a company that's investing in yours.

Hilbert is an Equal Opportunity Employer. We are committed to building a diverse team and an inclusive culture, and we welcome applicants of all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, gender identity or expression, sexual orientation, national origin, age, disability, veteran status, or any other status protected by applicable law.

Similar Jobs

Senior level
Artificial Intelligence • Hardware • Software • Semiconductor
Design, build, and operate CI/CD, Kubernetes-based platforms, deployment automation, and observability for engineering workflows. Improve reliability, performance, and scalability across cloud and on-prem environments, debug cross-boundary failures, perform root-cause analysis, and deliver durable platform software and self-service tooling.
Top Skills: Argo CdArtifact RepositoriesAWSCi/CdContainerized EnvironmentsCustom ResourcesHelmKubernetesKubernetes OperatorsLinuxMtlsPackage RegistriesPythonShellTerraformTls
Mid level
Artificial Intelligence • Hardware • Software • Semiconductor
Design, implement, and maintain Python frameworks and services that orchestrate distributed engineering workflows across machines and clusters. Build scheduling, execution, resource management, failure recovery, and test infrastructure. Define APIs and abstractions, reason about concurrency and distributed-systems behavior, debug complex multi-system issues, write automated tests and documentation, and partner with platform, CI, release, QA, and product teams to deliver scalable infrastructure.
Top Skills: AsyncioBuild SystemsCi SystemsCluster SchedulersConcurrent.FuturesContainersKubernetesMultiprocessingPytestPythonRelease InfrastructureRemote Execution Systems
12 Days Ago
Hybrid
San Jose, CA, USA
197K-246K Annually
Mid level
197K-246K Annually
Mid level
Fintech • Machine Learning • Payments • Software • Financial Services
Design, develop, deploy, and support foundational AI systems (foundation model training, LLM inference, similarity search, guardrails, evaluation, observability). Optimize training/inference for scalability, cost, latency, and throughput. Partner cross-functionally to deliver production AI services and shape the technical vision and roadmap for foundational AI at scale.
Top Skills: AWSAws UltraclustersAzureC#C++GoGCPHugging FaceJavaLlm InferenceNemo GuardrailsPythonPyTorchScalaSimilarity SearchVectordbs

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account