Bilingual Language Model Evaluator
# Bilingual Language Model Evaluator **mercor** · United States · `Remote` · `Contractor` 💼 **Уровень роли:** `Mid-level` 💰 **Компенсация:** `$15 – $20` 🕒 **Статус:** *Опубликовано: сегодня* · *Источник: Himalayas (JSON API)* --- ### Top Skills & Match 🎯 **Ключевой стек роли:** `[Bilingual-Evaluator]` `[Language-Model-Evaluator]` `[AI-Evaluator]` `[Data-Annotator]` `[Content-Evaluator]` `[Bilingual-AI-Evaluation]` `[Bilingual-Evaluation]` `[Language-Evaluator]` `[Bilingual-Content-Evaluator]` `[Language-Model-Evaluation]` `[Bilingual-Content-Evaluation]` `[Bilingual-Evaluation-Specialist]` `[Linguistic-Evaluator]` --- ### About the Role About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: Generalist - English & Urdu Type: Contract Compensation: $15–$20/hour Location: Remote Role Responsibilities - Conduct fact-checking using trusted public sources and external tools . - Generate high-quality human evaluation data by identifying response strengths, areas for improvement, and factual inaccuracies. - Assess reasoning quality, clarity, tone, and completeness of responses. - Ensure model responses align with expected conversational behavior and system guidelines. - Work independently and asynchronously to meet deadlines while improving AI model performance . Qualifications Must-Have - Bachelor's degree . - Native speaker in Urdu . - Significant experience using large language models (LLMs). - Excellent writing skills in English . - Strong attention to detail . - Background or experience in domains requiring structured analytical thinking . Preferred - Prior experience with RLHF, model evaluation, or data annotation work . - Experience writing or editing high-quality written content . - Experience comparing multiple outputs and making fine-grained qualitative judgments . Application Process (Takes 20–30 mins to complete) - Upload resume - AI interview based on your resume - Submit form Resources & Support - For details about the interview process and platform information, please check: - For any help or support, reach out to: PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. Originally posted on Himalayas
- Bilingual-Evaluator
- Language-Model-Evaluator
- AI-Evaluator
- Data-Annotator
- Content-Evaluator
- Bilingual-AI-Evaluation
- Bilingual-Evaluation
- Language-Evaluator
- Bilingual-Content-Evaluator
- Language-Model-Evaluation
- Bilingual-Content-Evaluation
- Bilingual-Evaluation-Specialist
- Linguistic-Evaluator
Наблюдалась 2026-10-02, впервые 2026-10-02, источник — Himalayas (JSON API).