Model Evaluation and Threat Research, Inc.
Platform Engineer, Application Security
We are a nonprofit research organization that develops scientific methods to assess AI capabilities, risks, and mitigations, with a specific focus on threats related to AI R&D automation and misalignment.
We believe it is robustly good for policymakers and civil society to have a clear understanding of risks from AI systems, and we are extremely excited to build a team of ambitious, excellent people to tackle one of the most important challenges of our time.
About the role
METR’s mission of enabling transparency and coordination about the risks of frontier AI requires a high degree of trust from frontier AI labs, governments, and the public. As misalignment incidents become more extreme and confidential information about models and frontier AI labs becomes more valuable, we expect to be under increasingly intense pressure from external actors and internal agents.
METR is looking to expand the security expertise on our platform team. This would span application security in our evaluation platform, sandboxing agents and evaluations, cloud platform security, networking, access control for both people and agents, and securing our development environments and workflows. This role will have a large engineering component: expect to write and review code, fix vulnerabilities, and design and develop secure systems.
What this role looks like
Securing a unique attack surface. METR's evaluation infrastructure runs frontier AI agents, including early checkpoints of unreleased models, executing untrusted, model-generated code at scale on multi-day tasks. You will design and build the isolation, networking, and permission boundaries that contain evaluated agents.
Remediation and hardening. You will find, triage, and fix vulnerabilities across our cloud infrastructure and access control systems.
Fixing known vulnerabilities in our codebase. This includes code reviews and resolving known vulnerabilities in our backlog.
Identity and access as a system. You will design and implement IAM policies, least-privilege access, and automated provisioning and access review for both people and agents.
Securing agentic systems. You will design and implement systems to monitor and control agents, making sure they can operate securely and with appropriate permissions and guardrails.
Required skills
7+ years of experience working in security engineering, software development, or an adjacent field.
Production software engineering. You have experience building and operating backend or infrastructure in practice.
Cloud and container security. You have deep familiarity with AWS (especially non-trivial IAM), Kubernetes, and infrastructure-as-code environments.
Vulnerability remediation. You have found and fixed security flaws in large production systems and can prioritize a remediation backlog.
Security fundamentals. Strong security knowledge across systems, networks, cloud, and identity, and a track record of applying it to real systems. Experienced in designing secure software and cloud architectures.
Code review. Reviewing and giving constructive security feedback on PRs, including from the FOSS community.
Nice to haves
Detection engineering at scale: Experience with detection pipelines (DataDog SIEM, AWS SecurityHub), writing and tuning detections, and threat hunting.
AI/LLM engineering: You build with AI: agent pipelines, LLM-powered tooling, automated workflows, and understand current limitations of those tools.
AI security research: Familiarity with agent control, hardware security, or red teaming AI systems themselves.
AWS: cloud-native software platforms
EKS
Lambda
ECS
IAM (in-depth)
CloudWatch
SecurityHub & GuardDuty
PostgreSQL: RLS, serverless Aurora
Pulumi: IaC
DataDog: SIEM
Okta: IdP
Google Workspace: IdP
Tailscale: networking
CrowdStrike Falcon: endpoint security
Ideally you have experience with a portion of these technologies:
Similar Jobs
What you need to know about the San Francisco Tech Scene
Key Facts About San Francisco Tech
- Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
- Major Tech Employers: Google, Apple, Salesforce, Meta
- Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
- Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
- Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
- Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine


.png)
