3Pillar Global Logo

3Pillar Global

AI Data Architect

Posted Yesterday
Be an Early Applicant
Remote
Hiring Remotely in United States
Expert/Leader
Remote
Hiring Remotely in United States
Expert/Leader
Design, build, govern, and own an enterprise AI data platform that ingests, transforms, stores, and serves data for AI consumers. Define multi-domain data models, data contracts, pipelines, vector and retrieval infrastructure, observability for agentic behavior, and evaluation frameworks. Establish architecture standards, CI/CD, and infrastructure-as-code to enable production AI/ML and LLM applications.
The summary above was generated by AI
3Pillar is an AI transformation partner on a mission to help enterprises build the AI-native products and intelligent agents that will define the next era of business. With teams across North America, Europe, Latin America, and Asia, we work with the most ambitious companies in financial services, healthcare, media, and technology — helping them move faster, modernize boldly, and compete on their own terms. Our HelixAI platform and Helix Pods delivery model put our engineers at the center of real agentic transformation — doing work that is open, portable, and built to last. We are building the future of enterprise AI.
 
AI Data Architect
 
We are looking for an AI Data Architect to design, build, govern, and evolve the single source of truth that powers every AI initiative in our organization.
This platform will serve as the foundational nervous system for conversational AI assistants, dashboard intelligence, autonomous AI agents, RAG-powered applications, predictive ML models, and any AI product we build today or in the future. The resource will architect the system, drive implementation, own the data contracts that agents and AI applications depend on, enforce security and access governance for both human and agent consumers, and continuously monitor and improve the accuracy and reliability of AI outputs that flow from this platform. 

Requirements:

    Architect and own the enterprise AI data platform — the unified, governed layer that ingests, transforms, stores, and serves all data consumed by AI systems across the organisation.
    Design multi-domain data models (lakehouse, data mesh, event-driven) that are structured from day one to serve AI workloads: clean lineage, versioned schemas, well-documented contracts, and low-latency serving APIs.
    Strong exposure to different Data architectures, data lake & data warehouse
    Define tools & technologies to develop automated data pipelines, write ETL processes, develop dashboard & report and create insights

Responsibilities

    Technical Skills
    Primary Skills:
     Python, SQL, Snowflake/Databricks, AWS (S3, Glue, EKS, Bedrock, Kinesis, Redshift), Docker, Kubernetes, Terraform, GitHub Actions, LangChain, LlamaIndex, LLM APIs (OpenAI, AWS Bedrock, Claude, HuggingFace), (Pinecone, FAISS, ChromaDB, OpenSearch), knowledge graphs (Neo4j).
     
    Secondary Skills: MLflow, FastAPI, CI/CD pipelines, observability tooling (CloudWatch, Grafana, or equivalent), data lineage and metadata management platforms.
     
    • 15+ years of hands-on data engineering and architecture experience, alongside building production AI/ML and LLM-era data infrastructure.
    • Strong Experience with either Databricks or Snowflake; experience with both is desirable.
    • Strong data architecture patterns & principles, ability to design secure & scalable data lakes, data warehouse, data hubs, and other event-driven architectures
    • Expertise in designing and writing ETL processes in Python / Java / Scala
    • Own the full data stack: real-time streaming (Kafka, Spark Structured Streaming), batch processing (Databricks, PySpark, Delta Lake), cloud storage and compute (AWS, Azure), and data quality /metadata management.
    • Drive modernisation of legacy pipelines (on-prem ETL, batch DWH) to cloud-native, AI-ready architectures with measurable improvements in cost, latency, and delivery velocity.
    • Proven experience designing enterprise-scale AI data platforms that serve multiple AI consumers —not just one application or pipeline.
    • Hands-on experience with vector stores, semantic models, knowledge graphs, and retrieval infrastructure in production environments.
    • Working knowledge of LLMOps: model serving pipelines, MLflow, CI/CD for AI, automated evaluation, and production monitoring.
       
       

AI Experience

    RAG, Vector & Retrieval Infrastructure
    Design the retrieval infrastructure that powers RAG-based AI applications: embedding pipelines, vector stores (Pinecone, FAISS, ChromaDB, OpenSearch), chunking strategies, and hybrid retrieval layers combining semantic search with structured queries.
     
    Agentic Behaviour Observability & Output Accuracy
    Own the observability stack for AI agent behaviour: instrument agents to capture inputs, retrieved context, tool calls, reasoning traces, and outputs — creating a complete audit trail of every agentic action driven by platform data.
    Design and operate evaluation frameworks that continuously measure AI output quality: factual accuracy, context faithfulness, retrieval relevance, hallucination rates, and task completion success— across all AI consumers of the platform.
     
    Architecture Standards & Engineering Enablement
    Define and maintain the reference architecture for the AI data platform — documenting design patterns, data contracts, integration standards, and decision records (ADRs) that all engineering teams follow.
    Establish data engineering standards: pipeline testing frameworks, code review practices, CI/CD automation, infrastructure-as-code (Terraform), reusable component libraries, and observability instrumentation.

Benefits

  • Medical Insurance benefits as per company policy. 

  • Dental insurance as per company policy.

  • Vision insurance as per company policy.

  • Employer paid Disability, Life, and AD&D insurance

  • Unlimited PTO 

  • Paid parental leave 

  • 401K

  • Flexible work policy

  • 12 Paid Holidays

Similar Jobs

3 Days Ago
Remote or Hybrid
US
124K-280K Annually
Senior level
124K-280K Annually
Senior level
Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Lead design and implementation of healthcare data architecture and pipelines, influence stakeholders, mentor teams, ensure compliance (HIPAA/PHI), apply interoperability standards (HL7, FHIR), and drive GenAI and data-driven improvements across health system operations.
Top Skills: Emr Back-End DatabasesFhir R4Gen AiHealthcare Integration EnginesHl7 V2Interoperability StandardsMiddlewarePythonSQL
9 Days Ago
Remote or Hybrid
Expert/Leader
Expert/Leader
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Lead enterprise AI transformation engagements as a senior strategic and technical advisor. Advise executives on AI strategy, design end-to-end AI and data architectures (RAG, knowledge graphs, MCP), architect data catalog and metadata strategies, embed AI governance and responsible AI controls, and drive adoption, change management, and practice development across customers.
Top Skills: Agent-To-Agent (A2A)Agentic WorkflowsAi AgentsAi Control TowerApp EngineAWSBiCloud PlatformsCsmData CatalogData LineageData WarehouseFsmGenerative AiGCPHrsdItsmKnowledge GraphsMachine LearningMetadata ManagementModel Context Protocol (Mcp)Now AssistRdfRetrieval-Augmented Generation (Rag)ServicenowSparql
9 Days Ago
Remote or Hybrid
Expert/Leader
Expert/Leader
Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Lead enterprise AI strategy and transformation engagements, advise executives, design end-to-end AI and data architectures (including RAG, knowledge graphs, data catalogs), define governance and responsible AI practices, drive adoption and change management, and develop reusable practice IP and thought leadership to enable scalable enterprise AI programs.
Top Skills: Agent-To-Agent (A2A)Agentic WorkflowsAi AgentsAi Control TowerAws Machine LearningData CatalogGoogle Professional Machine Learning EngineerKnowledge GraphsMetadata ManagementModel Context Protocol (Mcp)Now AssistRdfRetrieval-Augmented Generation (Rag)ServicenowSparql

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account