Jobs · Quality Assurance

Language Model Evaluator - Fully Remote | Upto $20/hr Part-time

Mercor · United States · 1 mo ago
RemoteRemoteQuality Assurance$15–$20/hrPart-time

Role Responsibilities

  • Conduct fact-checking using trusted public sources and external tools.
  • Generate high-quality human evaluation data by identifying response strengths, areas for improvement, and factual inaccuracies.
  • Audit reasoning quality, clarity, tone, and completeness of responses.
  • Ensure model responses align with expected conversational behavior and system guidelines.
  • Work independently and asynchronously to meet deadlines while improving AI model performance.

Qualifications

  • Must have a Bachelor's degree.
  • Must be a native speaker in Punjabi.
  • Must have significant experience using large language models (LLMs).
  • Must have excellent writing skills in English.
  • Must have strong attention to detail.
  • Prior experience with RLHF, model evaluation, or data annotation work is preferred.
  • Preferred experience writing or editing high-quality written content.
  • Preferred experience comparing multiple outputs and making fine-grained qualitative judgments.

Similar jobs