Jobs · Quality Assurance

Bilingual Language Model Evaluator

Mercor · United States · 1 mo ago
RemoteRemoteQuality Assurance$15–$20/hrPart-time

Role Responsibilities

  • Conduct fact-checking using trusted public sources and external tools.
  • Generate high-quality human evaluation data by identifying response strengths, areas for improvement, and factual inaccuracies.
  • Audit reasoning quality, clarity, tone, and completeness of responses.
  • Ensure model responses align with expected conversational behavior and system guidelines.
  • Work independently and asynchronously to meet deadlines while improving AI model performance.

Qualifications

  • Must have a Bachelor's degree.
  • Must be a native speaker in Urdu.
  • Must have significant experience using large language models (LLMs).
  • Must have excellent writing skills in English.
  • Must have strong attention to detail.
  • Preferred qualifications include prior experience with RLHF, model evaluation, or data annotation work, experience writing or editing high-quality written content, and experience comparing multiple outputs and making fine-grained qualitative judgments.

Similar jobs