Dragos Logo

Dragos

Senior Engineering Support Engineer

Posted An Hour Ago
Remote
Hiring Remotely in United States
130K-130K Annually
Senior level
Remote
Hiring Remotely in United States
130K-130K Annually
Senior level
Own Tier 3 escalations and incident response for complex customer issues involving a cybersecurity platform. Investigate distributed systems, reproduce defects, coordinate resolutions with engineering and product teams, and build Datadog observability dashboards, monitors, and alerts. Create runbooks, troubleshooting guides, RCAs, and postmortems; mentor support staff; identify recurring failure patterns; communicate with customers; and participate in an application-layer on-call rotation.
The summary above was generated by AI

At Dragos, the mission is personal. The systems we protect deliver the water you drink, power your home, and keep the hospitals your community depends on running. Those critical infrastructure systems that power our civilization around the world are under attack every day by adversaries. When those systems fail, people are immediately at risk. We are the global leader in xOT cybersecurity, combining technology, threat intelligence, and expert services. The people here chose this work because they understand what is at stake. Here, you will find a remote-first mission-driven team across North America, Europe, the Middle East, and APAC built on authenticity, transparency, and trust. If safeguarding the systems that protect your family, friends, and community is the kind of work that matters to you, you are in the right place. 

About the Role: 

Dragos Engineering Support is the bridge between our customers and the engineering teams that build the Dragos platform. When something breaks in a customer environment — or before it does — Engineering Support is the team that investigates, diagnoses, and drives resolution. We own Tier 3 technical escalations, application-layer observability and reliability, and the feedback loop that turns production signals into better software.

As a Senior Engineering Support Engineer, you'll serve as a primary escalation point for complex customer issues on a team that spans deep SRE and platform expertise. You'll own deep-dive investigations into the Dragos application stack, collaborate with engineering and product teams to drive resolution, and help build the monitoring and observability practices that keep our platform healthy. You'll also mentor Tier 1/2 Customer Experience support staff, maintain the knowledge base that keeps the team effective, and participate in the on-call rotation that ensures we're always watching the signals that matter.

Responsibilities: 

  • Lead escalation triage and incident response for complex, multi-component customer support cases — owning investigation and resolution coordination from intake to close, running technical incident command, and coordinating with infrastructure and product teams as diagnosis requires.
  • Validate, reproduce, and document confirmed defects emerging from customer escalations — producing clear bug reports with supporting evidence, reproduction steps, and expected behavior.
  • Build and maintain application observability — develop and refine Datadog monitors, dashboards, and alerts that provide visibility into customer-facing SLOs, platform health, and the early warning signals that enable proactive response.
  • Own Customer Experience communication through the lifecycle of escalated incidents, ensuring accurate, timely updates and appropriate customer visibility into resolution progress.
  • Translate field experience into documentation — author and maintain troubleshooting guides, runbooks, and playbooks that Tier 1/2 CX support staff can act on independently.
  • Drive pattern recognition across the customer base — identify recurring failure modes, surface trends to product teams, and actively collaborate on permanent fixes rather than workarounds.
  • Participate in on-call rotation, including occasional weekend coverage, triaging application-layer alerts within published SLOs and executing documented remediations.

Qualifications: 

  • 3+ years of experience in technical support engineering, site reliability engineering, or a closely related customer-facing engineering role, with demonstrated ownership of complex issue resolution.
  • Demonstrated experience troubleshooting complex distributed systems in production environments — comfortable reading logs, tracing component interactions, and isolating root cause under pressure.
  • Strong Linux system administration skills — comfortable with processes, filesystem, networking, and diagnosing application behavior from first principles.
  • Experience with containerized application environments (Kubernetes and/or Docker) at a level sufficient to investigate pod health, examine logs, and understand deployment state.
  • Familiarity with observability platforms (Datadog or equivalent) and some experience building and maintaining monitors, dashboards, and alerts.
  • Experience supporting Elasticsearch, PostgreSQL, or similar database platforms in a production environment.
  • Strong written communication skills with a demonstrated ability to produce clear, accurate technical documentation — RCAs, runbooks, postmortems — for both internal and customer-facing audiences.
  • Comfort working with AI tools and assistants as part of day-to-day engineering workflow, including prompt engineering and AI-assisted triage or investigation.
  • Experience operating in a customer-facing or customer-adjacent engineering role, with the judgment to balance urgency and thoroughness on competing priorities under SLA pressure.
  • Experience supporting or securing software in a cybersecurity context — working with security products, operating in security-conscious environments, or supporting customers with security-driven requirements.
  • Ability to read and navigate application source code well enough to trace a bug, understand a service's behavior, or follow a data flow.
  • Direct experience with OT/ICS cybersecurity environments or familiarity with industrial control systems — the product domain is learnable, but adjacent experience accelerates ramp significantly, preferred.
  • Experience with AWS (EKS, RDS, EC2) or comparable cloud infrastructure at the application layer — understanding how infrastructure behaviors manifest as application symptoms, preferred.
  • Background in customer success, customer engineering, or professional services roles that required deep technical credibility alongside relationship management, preferred.

Compensation: 

  • Salary: $130,000
  • Competitive Equity Package  
  • Comprehensive Benefits Plan 

 

#LI-MM1 #LI-REMOTE   


Dragos is an Equal Opportunity Employer and considers applicants for employment without regard to race, color, religion, sex, orientation, national origin, age, disability, genetics, or any other basis forbidden under federal, state, or local laws. All new hires must pass a background check as a condition of employment.

Similar Jobs at Dragos

An Hour Ago
Remote
United States
175K-175K Annually
Senior level
175K-175K Annually
Senior level
Security • Cybersecurity
Lead partner marketing strategy and execution to drive sourced and influenced pipeline. Manage MDF, co-marketing, ABM-aligned partner programs, enablement assets, partner engagement, and measurement in collaboration with demand generation, field marketing, and partner teams.
Top Skills: HubspotPrm PlatformsSalesforce
3 Hours Ago
In-Office or Remote
Senior level
Senior level
Security • Cybersecurity
Enable Southeast Asian channel partners to qualify, position, demonstrate, size, and support Dragos OT cybersecurity solutions. Build partner certification and training programs, create technical enablement assets, support partner-sourced opportunities, coach sellers, provide technical escalations, and coordinate evaluation deployments. The role requires cybersecurity, pre-sales, solution architecture, network monitoring, industrial-sector, and partner enablement expertise, with regular regional travel up to 50%.
Top Skills: Industrial ProtocolsNetwork ArchitectureNetwork Collection DesignOt CybersecurityPassive Monitoring
Yesterday
In-Office or Remote
230K-230K Annually
Expert/Leader
230K-230K Annually
Expert/Leader
Security • Cybersecurity
Leads complex ICS/OT cybersecurity consulting engagements for critical infrastructure organizations. Responsibilities include architecture reviews, threat assessments, compromise investigations, tabletop exercises, executive briefings, strategic recommendations, engagement governance, research, methodology improvement, thought leadership, mentorship, and customer relationship development. The role advises engineering, operations, and executive stakeholders while influencing security services, detection content, technology, and training.
Top Skills: Dnp3Endpoint Visibility PlatformsHmisIcs/Ot CybersecurityIndustrial InstrumentationMdrModbusNetwork Visibility PlatformsPlcsRtusSIEMSoar

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account