You'll evaluate AI-generated medical outputs in Wolof-English, assessing them for clinical accuracy, appropriateness, and linguistic quality. This work is conducted in partnership with a leading LLM research lab developing next-generation medical AI systems Wolof-qualified candidates are especially scarce — early expressions of interest are strongly encouraged
KEY RESPONSIBILITIES Evaluate LLM-generated Wolof-English medical content using structured rubrics Assess outputs for clinical accuracy, fluency, cultural appropriateness, and safety Provide written feedback and corrections in Wolof and/or English Flag any content that is clinically unsafe, culturally inappropriate, or linguistically incoherent Participate in optional calibration sessions with co-evaluators IDEAL QUALIFICATIONS Qualified medical professional (doctor, clinical officer, nurse with clinical degree, or equivalent) Native Wolof speaker — this is an absolute requirement; Strong English proficiency (clinical reading level) Based in Senegal, The Gambia, Mauritania, or Wolof-speaking diaspora community Self-motivated with strong attention to detail WHAT SUCCESS LOOKS LIKE High-quality, consistent evaluations that meet research-grade standards Willingness to commit to the full evaluation batch over 4–6 weeks Engagement with calibration process to ensure inter-annotator reliability CONTRACT & PAYMENT TERMS
Paid per reviewed entry or hourly (confirmed at onboarding). Payment via international transfer.