SuperAnnotate Logo

SuperAnnotate

Research Engineer

Posted 11 Days Ago
Be an Early Applicant
Hybrid
San Francisco, CA, USA
180K-250K Annually
Mid level
Hybrid
San Francisco, CA, USA
180K-250K Annually
Mid level
Perform end-to-end research engineering: find and reproduce relevant papers, implement and validate methods, build MVPs (including environments/eval harnesses and annotation workflows), partner with technical and strategic leads, and convert research into datasets, pilots, publications, or customer deliverables.
The summary above was generated by AI
About SuperAnnotate

SuperAnnotate helps the world’s leading AI teams build responsible, next-generation models powered by high-quality human data. We’re a fast-growing Series B startup bridging the gap between advanced AI innovation and the data that drives it. Our global network of expert specialists, scalable managed operations, precise talent matching, and full project transparency ensure unmatched data quality at scale. Trusted by innovators like Databricks and ServiceNow - and backed by NVIDIA, Dell Technologies Capital, Databricks Ventures, Cox Enterprises, and Lionel Messi’s Play Time VC - SuperAnnotate is proud to be the top-ranked AI data company on G2 for multiple consecutive years, including 2025.

The Impact You'll Make

Our research team is expanding to keep pace with a wave of frontier-facing work: internal research streams, client engagements that require real ML depth, and emerging opportunities at the cutting edge of the field. As a Research Engineer, you'll take a research direction and run with it – finding the right papers, benchmarks, and prior work, reimplementing what's relevant, and building out the process to reproduce and improve on it internally.

You'll own initiatives end to end: partnering with strategic project and technical leads to scope the work, building MVPs to validate ideas (including through human annotation and agents), and turning that work into something concrete – a customer dataset, a pilot, an internal dataset that becomes a paper or blog post, or a joint publication with a partner. You won't be handed a fully specified task list; you'll be given a direction and the autonomy to turn it into a research plan.

This is a full-time, hybrid position based in San Francisco.

What You'll Do

  • Take a research direction and independently identify supporting resources – papers, benchmarks, blog posts – then implement or reimplement the relevant methods.
  • Build and own the process to reproduce prior work internally and identify ways to improve on it.
  • Own projects (for example, an RL/agentic environment build for a partner or a novel multimodal benchmark) end to end, including scoping, MVP implementation, and validation.
  • Partner with strategic project leads and technical leads to translate ambiguous requirements into a concrete, testable research plan.
  • Validate ideas through hands-on implementation, including annotating, evaluating, or sourcing data.
  • Turn research directions into tangible outputs – a paid customer dataset, a customer pilot, an internal dataset, or a paper/blog post for publication or conference presentation.
  • Bring an ML perspective to new opportunities — assessing technical feasibility of incoming requests and helping shape proposals where research depth is needed.

What You'll Bring

  • MS or PhD in ML, CS, or a related quantitative field – or equivalent demonstrated research experience (publications, significant open-source research work, industry research).
  • Real ML depth: you understand how models are trained and evaluated, not just how to call an API. You can read a paper, judge whether its claims hold, and reimplement the method.
  • Hands-on experience with at least one of: RL/agentic systems, AI/ML evaluation and benchmarking, or multimodal ML.
  • Strong Python and the engineering ability to build and ship your own experiments – eval harnesses, environments, infrastructure – without relying on a platform team.
  • High autonomy: you can turn an ambiguous direction into a concrete research plan and notice when something's off before being told.
  • Clear technical writing

Nice To Have

  • Publication track record (first-author preferred).
  • Experience with agent or multimodal benchmarks (OSWorld, MMMU, WebArena, SWE-bench, or similar) or building RL environments/gyms.
  • Familiarity with reward modeling, reward hacking, or verifier/judge reliability.
  • Familiarity with synthetic data generation or human-in-the-loop (HITL) workflows.
  • Experience with cloud infrastructure and containerized environments.
  • A deep RL background specifically.

Why SuperAnnotate

This is a rare opportunity to work at the intersection of frontier AI research and real production impact. You'll work on projects with frontier labs that move the needle on model performance, with your work feeding directly into the next generation of agent capabilities. You'll have the opportunity to implement projects that actually matter, publish research, and present at conferences – alongside a multidisciplinary, multinational team and collaborate with some of the most prominent labs and AI companies globally.


Only shortlisted candidates will be contacted for an interview!

Equal Opportunity

We are an equal-opportunity employer and value diversity at our company. At SuperAnnotate diversity means to us making an effort to reflect the many experiences and identities of the outside world, and treating each other with fairness and without bias. Every day we foster an environment where people of all backgrounds not only belong, but excel to succeed as a company and grow together. We offer equal opportunity regardless of sex, sexual orientation, national origin, color, race, age, marital status, disability, gender identity, veterans and more.

HQ

SuperAnnotate San Francisco, California, USA Office

San Francisco, CA, United States

Similar Jobs

3 Days Ago
Hybrid
128K-160K Annually
Mid level
128K-160K Annually
Mid level
Artificial Intelligence • Hardware • Software • Nanotechnology • Semiconductor • Quantum Computing • Defense
Lead research and development of agentic AI and multi-agent systems integrating LLMs, memory, planning, tool use, and GraphRAG-style retrieval. Develop graph machine learning and geometric deep learning pipelines, build knowledge-augmented AI (knowledge graphs, ontologies), and ensure trustworthy AI via XAI, V&V, robustness testing, and uncertainty quantification. Collaborate across teams, publish research, and support proposal development for mission-critical autonomous and decision-support applications.
Top Skills: Agent2Agent (A2A)AgentopsAutogenCypherDistributed InferenceGeometric Deep LearningGnnsGpu AccelerationGraphragKnowledge GraphsLanggraphLlmopsLpgModel Context Protocol (Mcp)Neo4JOntologiesPythonPyTorchRayRdfSglangSparkVllm
4 Days Ago
Remote or Hybrid
35 Locations
100K-145K Annually
Mid level
100K-145K Annually
Mid level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
Design, build, and maintain distributed cloud services to support advanced threat detection. Collaborate with security, data science, and engineering teams to integrate capabilities into the Falcon platform, ensure production reliability, and leverage AI/LLMs to improve detection and automation. Mentor peers and contribute to scalable, high-availability system architecture.
Top Skills: AWSAzureBambooCassandraDockerElasticsearchFalcon PlatformFlinkGCPGoGrafanaJavaJenkinsKafkaKubernetesLlmsPythonRedisRestful ApisScala
16 Days Ago
In-Office
Sunnyvale, CA, USA
207K-275K Annually
Senior level
207K-275K Annually
Senior level
Cloud • Information Technology • Machine Learning
Lead applied research to advance continuous learning for agents: design and evaluate LLM post‑training and RL methods, implement and deploy experiments at scale, validate research on customer tasks, optimize GPU/distributed training, and mentor engineers while driving cross‑functional technical direction.
Top Skills: CudaFastapiGpuKubernetesMegatronPostgresTemporal

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account