Clarium Logo

Clarium

AI Engineer - Data Intelligence

Reposted 16 Hours Ago
Remote
Hiring Remotely in US
150K-180K Annually
Junior
Remote
Hiring Remotely in US
150K-180K Annually
Junior
As an AI Engineer at Clarium, you will build and maintain data enrichment pipelines, design classification workflows, analyze datasets, and ensure data quality, primarily using Python and SQL, and work closely with senior engineers and data scientists.
The summary above was generated by AI

Why Clarium?

The healthcare industry overspends on its supply chain by over $25B each year, the result of fragmented data, inefficient workflows, and wasted supplies. Clarium is fixing that. Our AI-powered platform, Astra OS, gives hospitals end-to-end visibility into their supply chain operations, automating workflows and surfacing actionable insights so supply chain teams can focus on what matters most: patient care. We're trusted by some of the world's leading health systems, including Yale New Haven Health, Stanford, Geisinger, and Kaiser Permanente.

Founded in 2020, Clarium has raised $43M in total funding. Our Series A was led by Northzone, with participation from General Catalyst, AlleyCorp, Kaiser Permanente Ventures, Texas Medical Center Ventures, and 1984 Ventures.

The Opportunity

AI-powered platforms, like Clarium’s, deliver the highest impact when they are supported by high-quality data. As we scale to more health systems and deepen our offering of intelligent, data-driven workflows, the master data enrichment pipeline (the system that classifies and contextualizes every product flowing through a hospital's supply chain) has become a critical growth lever. We're investing in the team and infrastructure to make that layer faster, smarter, and more reliable.

You'll join the Data Products team, a small, unusually senior group responsible for the data assets, data science, and analytics that drive measurable value for our clients. Day-to-day, you'll build and own components of our enrichment pipeline: classification workflows, entity resolution systems, evaluation harnesses, and the production tooling that keeps it all running. You'll work closely with engineers and data scientists who've shipped real ML systems at scale, and your work will feed directly into decisions made by supply chain teams at some of the country's leading health systems.

A rare early-career opportunity to learn fast and own real work from day one. As the first junior hire on the team, you won't be buried under layers of abstraction. You'll work directly alongside people who've done this before, on problems that actually matter. Short feedback loops, real stakes, and the kind of hands-on growth that's hard to find this early in a career. It's the opportunity many of us wish we'd had starting out.

In This Role You Will

  • Build and maintain components of Clarium's master data enrichment pipeline, the system that classifies and enriches every product flowing through our platform

  • Design and own classification and entity resolution workflows that combine deterministic logic and LLMs for production data processing

  • Build and operate evaluation harnesses, label sets, and regression suites (we use Braintrust) to measure and improve pipeline quality with confidence

  • Write production Python and SQL; the majority of your time will be spent in code, not in configuration tools

  • Analyze complex datasets using statistics and ML to surface actionable insights and inform pipeline improvements

  • Proactively audit data for quality issues; find the problems no one else has noticed yet, diagnose root causes, and ship fixes

What You'll Bring

  • Strong Python skills and a track record of writing production code, not just scripts or notebooks

  • Strong SQL, including complex joins, window functions, performance tuning, and data modeling

  • Comfort working in ambiguous environments; you can scope a problem, make a plan, and execute without hand-holding

  • A genuine, non-negotiable commitment to data quality; you treat silent bugs as real failures

  • Ability to go deep on an unfamiliar domain and develop meaningful expertise over time

Nice to Have

  • Experience with LLM integrations, prompt evaluation, or classification at scale

  • Familiarity with eval frameworks such as Braintrust, Promptfoo, or equivalent

  • Prior work in healthcare, supply chain, or another domain where data quality has direct operational consequences

Skills & Tools You'll Use

Need to Know: Python · SQL · PostgreSQL · CI/CD · Production observability

Nice to Know: Temporal · Braintrust · Snowflake · AWS · Sigma

What You Get at Clarium

Target Base Salary Range: $150K - $180K

The base salary Clarium offers may vary depending upon the ultimate scope and responsibilities of the position and on the candidate’s job-related knowledge, skills, and experience. The total package will include equity, in addition to a full range of medical and/or other benefits, depending on the position offered. Pay and benefits are subject to change at any time, consistent with the terms of any applicable compensation or benefit plans.

Incentive Stock Options proportionate to your salary

Fully remote, with a NYC co-working space available; distributed team across multiple time zones with opportunities for in-person time

Unlimited PTO

Top-tier health, vision, and dental benefits

401K

The opportunity to build on a strong foundational team with deep data and engineering roots at a stage where your work genuinely shapes the product

Equal Opportunity Statement

Clarium is committed to promoting an inclusive work environment free of discrimination and harassment. We value a diverse and balanced team where everyone can belong.

Similar Jobs

3 Days Ago
Remote or Hybrid
120K-215K Annually
Senior level
120K-215K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design, build, and deploy AI/ML and generative AI solutions and full-stack applications. Develop frontend experiences, backend services, microservices, APIs, and cloud-native data integrations. Implement RAG, LLMs, vector DBs, dashboards, reporting, and self-service analytics to support quality, patient safety, and workflow automation. Provide technical leadership, mentor engineers, and apply CI/CD, testing, observability, security, and performance best practices.
Top Skills: AgentsApache SupersetAutomated TestingAWSAzureAzure Ai ServicesCi/CdDatabricksEvent-Driven ArchitectureGCPGenerative AiJavaJavaScriptLangchainLlmsMicroservicesMicrosoft FabricObservabilityOpenaiPower BIPrompt EngineeringPythonRagReactRest ApisSemantic KernelSnowflakeSpring BootSQLTypescriptVector Databases
12 Hours Ago
Remote or Hybrid
North Carolina, USA
137K-278K Annually
Senior level
137K-278K Annually
Senior level
Cloud • Information Technology • Internet of Things • Professional Services • Software
Design, build, and operate scalable data architectures and pipelines integrating AI/ML and LLMs. Ensure data quality, security, and compliance; optimize warehouse performance; collaborate with data scientists and stakeholders; lead end-to-end projects and implement AI agent orchestration for enterprise automation.
Top Skills: DbtEtl/Elt ToolingJavaKubeflowLangchainLanggraphLarge Language Models (Llms)MlopsPythonRetrieval-Augmented Generation (Rag)ScalaSemantic SearchSnowflakeSnowflake Intelligence/Co-WorkSql (Snowflake)TeradataTransformer ArchitecturesVector Databases
3 Days Ago
Remote
US
Senior level
Senior level
Biotech
Design, build, and productionize agentic LLM workflows to improve clinical data extraction quality and reliability. Integrate AI capabilities into production pipelines, ensure scalability and cost-effectiveness, develop evaluation/monitoring for safe deployment in regulated environments, and lead technical design and delivery with cross-team collaboration.
Top Skills: Agent-Based Ai WorkflowsAmazon S3Amazon SqsAws FargateAws LambdaCi/CdDockerEvent-Driven ArchitectureLinuxLlmsObservabilityPythonRest ApisRetrieval-Based Ai

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account