Weekday, Inc. Logo

Weekday, Inc.

STEM Researcher - Computational Fields

Posted 3 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in United States
60-90 Hourly
Junior
Remote
Hiring Remotely in United States
60-90 Hourly
Junior
Design and author research-oriented, multi-step benchmark tasks for frontier AI models; create reference solutions and evaluation standards using Python and notebooks; review AI outputs to identify methodological and analytical flaws; collaborate with AI researchers to refine evaluation methodologies.
The summary above was generated by AI

This role is for one of our clients

Compensation: $60-$90 per hour

Join a pioneering AI initiative focused on developing the next generation of evaluation benchmarks for frontier AI models. We are seeking researchers from computational STEM disciplines—as well as computationally intensive social sciences and humanities—to bring the rigor of real-world research into AI evaluation.

In this role, you will transform scientific methodologies such as experimental design, hypothesis testing, and data-driven analysis into sophisticated, multi-step benchmark tasks that challenge state-of-the-art AI systems. Working closely with AI researchers, you'll help uncover subtle reasoning errors and methodological flaws that only experienced researchers can identify.

This is a fully remote, full-time engagement requiring approximately 35 hours per week.


RequirementsKey Responsibilities
  • Design complex, research-oriented benchmark tasks inspired by real-world scientific workflows, including study design, experimentation, hypothesis testing, and data analysis.
  • Develop comprehensive reference solutions using Python, notebooks, and computational tools with the rigor expected in professional research.
  • Define clear evaluation standards that distinguish sound scientific reasoning from plausible but incorrect conclusions.
  • Review AI-generated solutions, identifying methodological weaknesses, analytical errors, and flawed reasoning that experienced researchers would recognize immediately.
  • Collaborate with AI researchers and fellow domain experts to improve benchmark quality, consistency, and scientific rigor.
  • Contribute to the continuous refinement of evaluation methodologies for advanced AI systems.
Required Qualifications
  • Master's degree, PhD, or equivalent practical experience in a STEM discipline, computational social science, computational humanities, or another research-intensive field involving programming and data analysis.
  • Minimum 1 year of experience in an active research role within academia, industry, government laboratories, or a similar research environment.
  • Demonstrated experience performing computational research involving Python, data analysis, simulation, modeling, machine learning, or scientific computing.
  • Strong understanding of experimental design, hypothesis testing, statistical analysis, and rigorous interpretation of research findings.
  • Working knowledge of Git, integrated development environments (IDEs), and notebook platforms such as Jupyter or Google Colab.
  • Experience with AI evaluation, benchmark development, AI training, or task authoring is preferred.
  • Excellent analytical thinking, attention to detail, creativity, and the ability to solve complex, open-ended problems independently.
  • Strong written communication skills for documenting technical methodologies and research findings.
  • Ability to commit approximately 35 hours per week on a consistent basis.
Preferred Qualifications
  • Experience designing reproducible computational experiments or research workflows.
  • Familiarity with machine learning, large language models, or AI-assisted research tools.
  • Background in benchmark design, scientific software development, or computational research infrastructure.
  • Experience mentoring researchers, reviewing scientific work, or contributing to peer-reviewed publications.
Why Join
  • Help shape how next-generation AI systems are evaluated using rigorous scientific methodologies.
  • Collaborate with leading AI researchers working on frontier models and advanced evaluation frameworks.
  • Apply your research expertise to improve AI reasoning, reliability, and scientific accuracy.
  • Contribute to impactful work that advances the quality and robustness of AI systems across multiple disciplines.
  • Enjoy the flexibility of a fully remote engagement while working on cutting-edge AI research initiatives.
Equal Opportunity

We are committed to fostering an inclusive and diverse environment where all qualified applicants receive equal consideration. Reasonable accommodations are available throughout the application and engagement process.

Contract & Engagement Details
  • Independent contractor engagement.
  • Fully remote with flexible working hours.
  • Expected commitment of approximately 35 hours per week.
  • Project duration may be extended, shortened, or concluded based on project requirements and individual performance.
  • Work does not require access to confidential or proprietary information from any current or former employer.
  • Payments are issued weekly based on approved work completed.
  • At this time, we are unable to support H1-B or STEM OPT candidates.

Similar Jobs

51 Minutes Ago
Remote or Hybrid
Chatsworth Lake Manor, CA, USA
23-31 Hourly
Entry level
23-31 Hourly
Entry level
Fintech • Financial Services
Serve as a frontline Personal Banker building customer relationships, supporting account openings, cash handling and teller activities, processing service requests and credit applications, and promoting banking products. Use digital tools, follow compliance and risk controls, coordinate referrals to specialists, and support branch growth through proactive outreach and teamwork.
2 Hours Ago
Remote or Hybrid
67K-138K Annually
Senior level
67K-138K Annually
Senior level
Artificial Intelligence • Fintech • Insurance • Marketing Tech • Software • Analytics
Provide analytics and insights to Enterprise Procurement by analyzing large datasets, developing dashboards, leading project workstreams, presenting recommendations to stakeholders, improving reporting infrastructure, and mentoring teammates. Identify expense-saving opportunities, assess tools/processes, and apply analytic methods and AI/GenAI where appropriate.
Top Skills: Ai/GenaiDatabricksDaxEmblemExcelPower BIPowerPointPythonSASSnowflakeSQL
4 Hours Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
140K-170K Annually
Senior level
140K-170K Annually
Senior level
Artificial Intelligence • Big Data • Logistics • Machine Learning • Software • Transportation
Sell FourKites SaaS supply chain and logistics solutions to new and existing Fortune 1000 accounts. Manage 15-25 accounts, exceed quota, develop strategic account plans, map solutions to customer SOPs, coordinate cross-functional GTM efforts, update Salesforce, and leverage internal AI tools to drive growth and expanded ARR.
Top Skills: Fourkites Ai ToolsLinkedin Sales NavigatorSaaSSalesforceZoominfo

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account