Berkeley Research Group Logo

Berkeley Research Group

AI Lab Infrastructure Engineer

Posted 5 Days Ago
Remote
Hiring Remotely in USA
Senior level
Remote
Hiring Remotely in USA
Senior level
Design and implement a virtual access layer to enable remote use of a physical AI lab, building scalable, secure infrastructure to process 100,000+ documents/day with LLMs, optimize GPU/LLM throughput, provide customizable interfaces, implement authentication/access controls, and lead projects from design to production.
The summary above was generated by AI
We do Consulting Differently

Berkeley Research Group's Ai Department is seeking an AI Infrastructure Engineer to lead the development of our Virtual Ai Lab initiative. Following the successful completion of Phase 01 (physical Ai Lab build-out), this role will focus on creating a virtual access layer that makes our high-performance Ai Lab remotely accessible to teams across BRG. The ideal candidate will design and implement scalable infrastructure to support processing 100,000+ documents daily using state-of-the-art LLMs from OpenAI and Anthropic.

About the Role

As an AI Infrastructure Engineer, you will architect and build the virtual access interface for our physical Ai Lab, ensuring secure, scalable, and efficient remote processing capabilities. You will lead the design and implementation of infrastructure that allows BRG teams to leverage our Ai Lab's computational power remotely, while maintaining performance standards for large-scale document processing. Key responsibilities include developing customizable interfaces for different BRG groups, implementing secure access controls, and ensuring optimal resource allocation for concurrent users processing massive datasets through LLMs.

Key Responsibilities

  • Design and implement a virtual access layer for the physical Ai Lab infrastructure
  • Build scalable remote processing capabilities supporting 100,000+ documents per day
  • Create customizable, expandable interfaces for different BRG business units
  • Optimize infrastructure for maximum LLM token throughput (OpenAI/Anthropic)
  • Implement secure authentication and access management systems
  • Ensure high availability and fault tolerance for mission-critical AI workloads
  • Lead infrastructure projects from conception to production deployment

Required:

  • Bachelor's degree in Computer Science, Information Technology, or a related field
  • Minimum six to eight (6-8) years of hands-on experience designing, deploying, and managing scalable cloud infrastructure
  • Strong experience with Infrastructure as Code (IaC) tools and methodologies
  • Experience designing, implementing, and maintaining scalable, secure, and cost-efficient cloud/on-prem solutions
  • Proven ability to manage and lead projects to deliver high-quality, replicable solutions
  • Proficiency in VCS (Git/GitHub), modern coding languages (Python, .NET, Java, etc.), Software Development Life Cycle, and CI/CD practices
  • Experience with API design and implementation for distributed systems
  • Knowledge of GPU infrastructure and optimization for AI workloads
  • Hands-on experience with AWS Services including:
    • EC2/Lambda (apps/functions)
    • SageMaker (ML)
    • S3 (file management)
    • Fargate/ECS/EKS (containerization)
    • CDK/Terraform (IaC)
    • Cost Explorer/Budgets

Preferred:

  • Experience with LLM deployment and optimization (OpenAI, Anthropic, etc.)
  • Background in building AI/ML infrastructure and platforms
  • Experience with virtual desktop infrastructure (VDI) or remote access solutions
  • Knowledge of distributed computing and job scheduling systems
  • AWS certifications (Solutions Architect, Machine Learning, or similar)
  • Experience with cost management and optimization strategies in the cloud
  • Familiarity with security best practices for AI systems and data handling

About BRG
 
BRG combines world-leading academic credentials with world-tested business expertise purpose-built for agility and connectivity, which sets us apart—and gets you ahead.

At BRG, our top-tier professionals include specialist consultants, industry experts, renowned academics, and leading-edge data scientists. Together, they bring a diversity of proven real-world experience to economics, disputes, and investigations; corporate finance; and performance improvement services that address the most complex challenges for organizations across the globe.

Our unique structure nurtures the interdisciplinary relationships that give us the edge, laying the groundwork for more informed insights and more original, incisive thinking from diverse perspectives that, when paired with our global reach and resources, make us uniquely capable to address our clients’ challenges. We get results because we know how to apply our thinking to your world.

At BRG, we don’t just show you what’s possible. We’re built to help you make it happen.  

BRG is proud to be an Equal Opportunity Employer. Our hiring practices provide equal opportunity for employment without regard to race, religion, color, sex, gender, national origin, age, United States military veteran status, ancestry, sexual orientation, marital status, family structure, medical condition including genetic characteristics or information, veteran status, or mental or physical disability so long as the essential functions of the job can be performed with or without reasonable accommodation, or any other protected category under federal, state, or local law.

HQ

Berkeley Research Group Emeryville, California, USA Office

2200 Powell Street, Suite 1200, Emeryville, CA, United States, 94608

Berkeley Research Group San Francisco, California, USA Office

San Francisco, United States

Similar Jobs

5 Minutes Ago
Easy Apply
Remote or Hybrid
San Jose, CA, USA
Easy Apply
102K-128K Annually
Junior
102K-128K Annually
Junior
Cloud • Information Technology • Security • Software • Cybersecurity
Drive automation-first reliability for a global, multi-cloud platform: build scalable infra (AWS/GCP/bare-metal), write automation (Python/Go), implement observability (Prometheus/Grafana/OpenTelemetry), lead incident response/on-call, define SLIs/SLOs, and partner on operability reviews and post-incident analysis.
Top Skills: AnsibleAWSAzureBgpC/C++DnsGCPGoGrafanaGreHaproxyHelmIpsecItilLinuxOpentelemetryPrometheusPythonRhelTemporalTerraform
11 Minutes Ago
Remote or Hybrid
118K-201K Annually
Senior level
118K-201K Annually
Senior level
Aerospace • Hardware • Information Technology • Security • Software • Cybersecurity • Defense
Lead supplier quality for Printed Wiring Boards: audit suppliers, perform source and first-article inspections, drive root-cause analysis and corrective actions, implement process improvements, and ensure compliance with PWB and aerospace standards to deliver first-time quality.
Top Skills: ApqpAs9100As9102Asme Y14.5Asme Y15.1Black BeltControl PlanFirst Article InspectionGreen BeltIpc-6012Ipc-6013Ipc-6018Ipc-A-600Ipc-A-610Ipc-Tm-650Lean Six SigmaMil-Prf-31032Mil-Prf-38534Mil-Prf-55110Mil-Std-883PfmeaPpapSource Inspection
11 Minutes Ago
Remote or Hybrid
District of Columbia, USA
127K-215K Annually
Mid level
127K-215K Annually
Mid level
Aerospace • Hardware • Information Technology • Security • Software • Cybersecurity • Defense
Support and maintain complex applications and infrastructure for a government customer: monitor and triage events, troubleshoot Linux/Windows servers, deploy and integrate software (AWS, CloudFormation, RDS), use Salt for configuration management, work with databases (Oracle, MongoDB, PostgreSQL, MySQL), write SOPs, manage security groups, and support after-hours deployments. Requires strong communication and collaboration with developers and vendors.
Top Skills: AWSCloudFormationElasticsearchJavaScriptLinuxMongoDBMySQLOraclePostgresPythonRdsSaltstackWindows Server

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account