Bilingual Language Model Evaluator
Mercor · United States · 1 mo ago
RemoteRemoteQuality Assurance$15–$20/hrPart-time
Role Responsibilities
- Conduct fact-checking using trusted public sources and external tools.
- Generate high-quality human evaluation data by identifying response strengths, areas for improvement, and factual inaccuracies.
- Audit reasoning quality, clarity, tone, and completeness of responses.
- Ensure model responses align with expected conversational behavior and system guidelines.
- Work independently and asynchronously to meet deadlines while improving AI model performance.
Qualifications
- Must have a Bachelor's degree.
- Must be a native speaker in Urdu.
- Must have significant experience using large language models (LLMs).
- Must have excellent writing skills in English.
- Must have strong attention to detail.
- Preferred qualifications include prior experience with RLHF, model evaluation, or data annotation work, experience writing or editing high-quality written content, and experience comparing multiple outputs and making fine-grained qualitative judgments.