Baseten Logo

Baseten

AI Engineer

Posted 6 Days Ago
Hybrid
San Francisco, CA, USA
220K-260K Annually
Senior level
Hybrid
San Francisco, CA, USA
220K-260K Annually
Senior level
Build and ship agentic AI product experiences, internal automation, and reliable production AI systems. Partner with research engineers to turn internal workflows into customer-facing features, designing harnesses, execution flows, guardrails, and evaluations. Work end to end across APIs, backend systems, agent orchestration, and frontend interfaces. Identify manual processes suitable for AI automation, resolve customer issues, and operate autonomously in a fast-moving environment.
The summary above was generated by AI

ABOUT BASETEN

Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products.

THE ROLE:

Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster.

You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk.

EXAMPLE INITIATIVES:

Take a look at these blog posts written by members of our team:

  • Baseten Training: an autoresearch substrate

  • Introducing Baseten Loops

  • Harnesses are everything. Here's how to optimize yours.

  • Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten

RESPONSIBILITIES:

  • Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA.

  • Design the harnesses, execution flows, and guardrails that make AI systems reliable in production.

  • Build internal automation and AI tooling that measurably increases the velocity of engineering, research, and go-to-market teams.

  • Partner with research engineers to scope product opportunities out of internal research workflows and turn them into customer-facing features.

  • Define and instrument evals so you know whether a change actually improved output quality.

  • Work throughout the stack (API layer, backend, agent orchestration, frontend) to implement features end to end.

  • Use Baseten's own training and inference products yourself to develop intuition around customer workflows.

  • Identify where AI can replace manual process across the company and build the thing rather than write the proposal.

  • Fix bugs and resolve customer issues with urgency.

REQUIREMENTS:

  • 5+ years of experience building and shipping software applications.

  • Demonstrated experience building AI or LLM-powered products, agents, or agentic workflows that real users depend on.

  • Strong software engineering fundamentals and the ability to clear a real technical bar, not just prompt well.

  • Ability to build accurate mental models of how systems work under the hood, including the models and harnesses you're building on.

  • Proficiency in Python, with fluency in at least one other language.

  • Comfort working autonomously in a fast-moving environment with limited structure.

  • Ability to move between customer-facing product work and internal tooling and automation.

  • Strong communication skills, with the ability to bridge technical depth and business needs.

NICE TO HAVE:

  • Experience as a founding engineer or early employee at a startup.

  • Experience building evals, agent observability, or tooling for non-deterministic systems.

  • Familiarity with agent frameworks and harnesses (LangChain, Claude Code, Codex, OpenCode, MCP).

  • Experience with model development methods like supervised fine-tuning, reinforcement learning, synthetic data generation, LoRA, and full fine-tunes.

  • Frontend fluency.

BENEFITS

  • Competitive compensation, including meaningful equity.

  • 100% coverage of medical, dental, and vision insurance for employee and dependents

  • Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)

  • Paid parental leave

  • Fertility and family-building stipend through Carrot

  • Company-facilitated 401(k)

  • Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.

Apply now to embark on a rewarding journey in shaping the future of AI! If you are a motivated individual with a passion for machine learning and a desire to be part of a collaborative and forward-thinking team, we would love to hear from you.

At Baseten, we are committed to fostering a diverse and inclusive workplace. We provide equal employment opportunities to all employees and applicants without regard to race, color, religion, gender, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, or veteran status.

We are an Equal Opportunity Employer and will consider qualified applicants with criminal histories in a manner consistent with applicable law (by example, the requirements of the San Francisco Fair Chance Ordinance, where applicable).

HQ

Baseten San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

Yesterday
Hybrid
San Jose, CA, USA
186K-282K Annually
Senior level
186K-282K Annually
Senior level
Artificial Intelligence • Fintech • Software
Own FloQast’s multi-region AI platform infrastructure across AWS, including Bedrock model runtimes, sandboxed code execution, Terraform, observability, cost controls, CI/CD, security, and reliability. Support Transform, AI Matching, and AutoBuilder by managing scaling, throttling, tenant isolation, journey-based SLOs, model and prompt delivery gates, and AI-specific on-call processes across US, EU, and AU regions.
Top Skills: SparkAws AlbAws BedrockAws Bedrock AgentcoreAws EcsAws FargateAws IamAws LambdaAws NlbAws S3Aws SqsAws VpcDockerEmrFinopsGithub ActionsGrafanaHarnessMongoDBNode.jsNxOpentelemetryPostgresPrometheusPythonSnowflakeTerraformTruefoundryTypescript
Yesterday
Easy Apply
Hybrid
San Francisco, CA, USA
Easy Apply
240K-305K Annually
Senior level
240K-305K Annually
Senior level
Fintech • HR Tech
Own Gusto’s agentic coding platform, including autonomous coding workflows, AI code review, orchestration, security posture, sandboxing, permissions, infrastructure contracts, and AI spending controls. Lead and grow a senior engineering team, recruit and calibrate talent, set the product and technical roadmap, manage ambiguity across multiple domains, and advise executive leadership on AI-native software development strategy.
Top Skills: Ci/CdClaude CodeCliCloud InfrastructureCodexCursorPermission ScopingRuntime IsolationSandboxingSecrets ManagementSlack
3 Days Ago
Remote or Hybrid
San Francisco, CA, USA
245K-335K Annually
Senior level
245K-335K Annually
Senior level
Fintech • Machine Learning • Payments • Software • Financial Services
Design, develop, test, deploy, and support production AI systems, including foundation model training, LLM inference, similarity search, guardrails, evaluation, governance, experimentation, and observability. Develop state-of-the-art optimization techniques to improve scalability, cost, latency, throughput, and hardware utilization. Shape the long-term technical roadmap for foundational AI platforms while collaborating with engineers, research scientists, program managers, and product managers. Lead and mentor engineering teams and influence senior cross-functional stakeholders.
Top Skills: AWSAws UltraclustersAzureC#C++GoGCPHugging FaceJavaNemo GuardrailsPythonPyTorchScalaVectordbs

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account