Metasys Logo

Metasys

DevOps Engineer Internship

Reposted Yesterday
Remote
Hiring Remotely in United States
Internship
Remote
Hiring Remotely in United States
Internship
Build and maintain automated infrastructure and CI/CD pipelines using Terraform, Docker, Traefik, and Makefile. Implement observability (Prometheus, Grafana, Loki, Tempo, OpenTelemetry), backups (pgBackRest/Postgres15), and cloud/Linux administration. Automate deployment and monitoring for AI agent services and support monorepo workflow with SRE and DevSecOps teams.
The summary above was generated by AI
Overview: Infrastructure Automation and CI/CD

The DevOps Engineer is responsible for automating, streamlining, and maintaining the infrastructure and deployment pipelines for our entire integrated platform. You'll ensure rapid, reliable, and consistent delivery of our e-commerce storefront, internal supply chain tools (MES, WMS, OMS), and cutting-edge AI agent services, primarily utilizing Infrastructure-as-Code (IaC) and robust CI/CD practices.

Internship Details

Duration: 3 months
Start Date: Immediate
Location: Remote
Stipend: None initially. Based on your first-quarter performance, you may be offered a paid full-time opportunity, or even be absorbed directly by the client as an FTE.

Key Responsibilities & Core Projects

You will build and maintain the fully automated platform that underpins our entire tech stack.

  • Infrastructure-as-Code (IaC): Design, implement, and manage infrastructure provisioning across all environments using Terraform for our Oracle Cloud Free VMs (or equivalent cloud resources). Ensure infrastructure is auditable, repeatable, and secure.

  • CI/CD Pipeline Management: Set up and maintain the Continuous Integration and Continuous Deployment (CI/CD) pipelines, primarily driven by Makefile and automated testing, for the Node.js/NestJS modular monolith and Next.js frontend applications.

  • Containerization & Orchestration: Manage application containerization using Docker. Define deployment strategies, service discovery, and traffic routing using Traefik for our containerized services.

  • Observability Implementation: Implement, manage, and optimize the comprehensive logging, monitoring, and alerting system using our selected stack: Prometheus, Grafana, Loki, Tempo, and OpenTelemetry. Ensure end-to-end tracing is functional across the complex business flow (MES → WMS → OMS).

  • Resilience & Backups: Collaborate with the SRE team to implement high-availability features and maintain automated backup solutions, including pgBackRest for our PostgreSQL 15 database.

  • Workflow: Maintain the Monorepo structure for streamlined code management and deployment separation across applications (web / admin / API) and domain packages.

Required Technologies & Tools

Candidates must possess mandatory expertise in our core infrastructure and automation stack:

  • Infrastructure-as-Code: Expert proficiency in Terraform.

  • Containerization: Expert proficiency in Docker and deployment strategies (e.g., Traefik, orchestration concepts).

  • CI/CD: Hands-on experience building and maintaining complex pipelines (Makefile, Jenkins/GitHub Actions/GitLab CI concepts).

  • Observability: Strong implementation experience with Prometheus, Grafana, Loki, and OpenTelemetry.

  • Cloud & Linux: Experience with Linux administration and managing cloud resources (Oracle Cloud or equivalent).

AI Agent Focus

You will ensure the scalable and monitored deployment of the AI layer.

  • Deployment Automation: Automate the packaging and deployment pipelines for resource-intensive AI agent services and LLM fine-tuning environments.

  • Resource Monitoring: Set up specific monitoring and alerts to track the performance, resource consumption, and cost of the AI agent compute demands.

Success Metrics & Career Path

Performance will be measured by:

  • Deployment Frequency: Reduction in lead time and increased frequency of stable deployments.

  • Infrastructure Stability: Reliability of provisioned infrastructure (minimal unplanned downtime).

  • Observability Coverage: Completeness and reliability of monitoring, logging, and tracing across all production services.

Mentorship Structure: Reports to the Solution Architect or Head of Technology, working closely with the SRE, DevSecOps, and Backend engineering teams to build a robust platform.

Similar Jobs

An Hour Ago
Remote or Hybrid
USA
91K-203K Annually
Senior level
91K-203K Annually
Senior level
Machine Learning • Payments • Security • Software • Financial Services
Leads complex, multi-workstream product initiatives as a principal Product Owner. Owns product vision, strategy, backlog integrity, prioritization, and Scrum team alignment. Coordinates Product, Engineering, Design, Delivery, and other stakeholders; improves execution discipline, transparency, and outcome tracking. Stabilizes programs with fragmented ownership or elevated delivery risk while ensuring customer and business requirements are addressed. Requires substantial Product Owner or Senior Product Manager experience, preferably in payments or financial services.
Top Skills: Agile DevelopmentData VisualizationScrumUx Design
3 Hours Ago
Remote or Hybrid
MI, USA
Senior level
Senior level
Software • Analytics • Hospitality
Leads revenue recognition, billing, accounts receivable, collections, audit readiness, and financial control operations. Ensures ASC 606 and GAAP compliance, manages billing platforms and ERP integrations, drives automation and AI adoption, develops scalable processes, partners cross-functionally, reports financial metrics, and supervises accounting staff. Requires extensive revenue accounting and billing experience in SaaS or software, management expertise, and strong technical accounting knowledge.
Top Skills: Ai-Enabled Accounting Automation ToolsBilling PlatformsErp SystemsExcelMicrosoft OutlookMicrosoft WordNetSuiteOracleSAP
3 Hours Ago
Remote or Hybrid
MN, USA
Senior level
Senior level
Software • Analytics • Hospitality
Leads revenue recognition, billing, accounts receivable, collections, audit readiness, and financial controls. Owns ASC 606 compliance, billing systems, process automation, AI-enabled accounting improvements, reporting, and cross-functional alignment. Manages and mentors the accounting team while supporting scalable revenue operations and public-company readiness.
Top Skills: Ai-Enabled Automation ToolsAsc 606GaapExcelMicrosoft OutlookMicrosoft WordNetSuiteOracleSAP

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account