Weekday, Inc. Logo

Weekday, Inc.

AI Rater Guidelines Writer (Linguist / Instructional Designer)

Reposted 22 Days Ago
Remote
Hiring Remotely in United States
45-65 Hourly
Mid level
Remote
Hiring Remotely in United States
45-65 Hourly
Mid level
Create clear, structured rater guidelines, evaluation rubrics, and reviewer instructions for human evaluators across domains. Review and revise documentation to remove ambiguity, design scoring frameworks, handle edge cases, and collaborate with researchers and SMEs to ensure consistent, high-quality AI evaluations.
The summary above was generated by AI

This role is for one of our clients

Compensation: $45-$65 per hour

Join a cutting-edge AI initiative focused on building the next generation of foundational AI models. We are seeking experienced Linguists, Instructional Designers, Technical Writers, and AI Content Specialists who excel at transforming complex, ambiguous program requirements into clear, structured, and actionable guidance for human evaluators.

In this role, you will play a critical part in improving AI training quality by creating precise rating guidelines, evaluation rubrics, and reviewer instructions across diverse domains including finance, retail, insurance, legal, and sports. Your expertise will help ensure consistency, accuracy, and reliability in human evaluations that directly shape advanced AI systems.

This is a full-time, fully remote engagement requiring a commitment of approximately 35 hours per week during standard weekdays.


RequirementsKey Responsibilities
  • Translate complex and ambiguous program requirements into clear, concise, and easy-to-follow rater guidelines.
  • Design comprehensive evaluation rubrics and scoring frameworks that enable consistent decision-making across a variety of AI tasks.
  • Review existing documentation to identify ambiguity, inconsistencies, contradictions, and coverage gaps, then revise guidelines to improve clarity and usability.
  • Convert subject-matter requirements from domains such as finance, retail, insurance, legal, and sports into structured evaluation instructions that non-domain raters can apply confidently.
  • Collaborate with research teams, program managers, and subject matter experts to maintain consistency across multiple guideline sets.
  • Develop documentation that addresses edge cases, exceptions, and complex evaluation scenarios while minimizing reviewer escalation.
Required Qualifications
  • Minimum 3 years of professional experience in Linguistics, Instructional Design, Technical Writing, AI Content Development, or a closely related field.
  • Direct experience creating, refining, or maintaining evaluation guidelines, rating rubrics, or reviewer instructions within Generative AI, RLHF, human evaluation, or AI data annotation environments.
  • Demonstrated ability to work across multiple subject areas and convert domain-specific knowledge into clear, structured documentation.
  • Proven experience resolving ambiguity and improving written specifications with measurable before-and-after improvements.
  • Strong analytical thinking with exceptional attention to detail.
  • Demonstrated career progression and professional growth.
  • Ability to commit reliably to 35+ hours per week during weekdays.
  • Outstanding written communication skills with the ability to explain nuanced concepts in a precise, consistent, and easy-to-understand manner.
Preferred Qualifications
  • Experience supporting AI model training, evaluation, RLHF, or large-scale annotation programs.
  • Background working with cross-functional teams including researchers, engineers, product managers, and subject matter experts.
  • Familiarity with structured documentation standards, quality assurance methodologies, and guideline governance.
  • Experience designing documentation that supports scalable, high-quality human evaluation processes.
Why Join
  • Help define how human evaluators assess next-generation AI systems.
  • Influence the quality and consistency of AI training data across multiple industries.
  • Collaborate with multidisciplinary experts working on advanced Generative AI initiatives.
  • Solve complex language and reasoning challenges while improving AI evaluation standards at scale.
  • Enjoy the flexibility of a fully remote engagement while contributing to high-impact AI research.
Equal Opportunity

We are committed to creating an inclusive workplace and welcome applications from qualified professionals of all backgrounds. Reasonable accommodations are available throughout the application and engagement process.

Contract & Engagement Details
  • Independent contractor engagement.
  • Fully remote with flexible working arrangements.
  • Expected commitment of approximately 35 hours per week during weekdays.
  • Project duration may vary depending on business requirements and individual performance.
  • Work does not require access to confidential or proprietary information from any current or former employer.
  • Payments are issued weekly based on approved work completed.
  • At this time, we are unable to support H1-B or STEM OPT candidates.

Similar Jobs

12 Minutes Ago
Easy Apply
Remote or Hybrid
United States
Easy Apply
134K-178K Annually
Entry level
134K-178K Annually
Entry level
eCommerce • Healthtech • Kids + Family • Retail • Social Media
Leads payer onboarding, transitions, ongoing payer operations, licensing, credentialing, audits, site inspections, and regulatory compliance for Babylist Health. Builds scalable processes, monitors payer performance, manages commercial and Medicaid operations, program-manages new business launches, and partners with billing, legal, product, contracting, and customer support. Manages and develops two direct reports while establishing the operational foundation for expanding clinical services.
Top Skills: Artificial IntelligenceElectronic Funds Transfer (Eft)MedicaidRevenue Cycle Management (Rcm)
12 Minutes Ago
Easy Apply
Remote
United States
Easy Apply
Entry level
Entry level
Cloud • Security • Software • Cybersecurity • Automation
Sources and engages passive candidates for sales roles across the AMER region, primarily Account Executives. Builds talent pipelines, maps competitive markets, screens candidates, partners with recruiters and hiring managers, maintains applicant tracking data, and improves outbound sourcing strategies and candidate experience in a fully remote environment.
Top Skills: AIGitlabGreenhouseLinkedInLinkedin Recruiter
13 Minutes Ago
Remote or Hybrid
136K-245K Annually
Expert/Leader
136K-245K Annually
Expert/Leader
Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Manage enterprise KYC due diligence policy and strategy across Block’s business lines. Develop and maintain KYC standards, assess regulatory changes, guide implementation, define AI/ML risk-segmentation and identity-verification controls, monitor program performance, advise stakeholders, support audits and examinations, and remediate compliance risks including model bias, drift, false positives, false negatives, and explainability gaps.
Top Skills: Ai/Ml

What you need to know about the San Francisco Tech Scene

San Francisco and the surrounding Bay Area attracts more startup funding than any other region in the world. Home to Stanford University and UC Berkeley, leading VC firms and several of the world’s most valuable companies, the Bay Area is the place to go for anyone looking to make it big in the tech industry. That said, San Francisco has a lot to offer beyond technology thanks to a thriving art and music scene, excellent food and a short drive to several of the country’s most beautiful recreational areas.

Key Facts About San Francisco Tech

  • Number of Tech Workers: 365,500; 13.9% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Google, Apple, Salesforce, Meta
  • Key Industries: Artificial intelligence, cloud computing, fintech, consumer technology, software
  • Funding Landscape: $50.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Sequoia Capital, Andreessen Horowitz, Bessemer Venture Partners, Greylock Partners, Khosla Ventures, Kleiner Perkins
  • Research Centers and Universities: Stanford University; University of California, Berkeley; University of San Francisco; Santa Clara University; Ames Research Center; Center for AI Safety; California Institute for Regenerative Medicine

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account