In this hourly, remote contractor role, you will work as a Data Scientist Quality Assurance Lead (QAL) to oversee quality, consistency, and trainer performance across data science AI training projects. You will review AI-generated data science content and trainer/QA work, evaluate output quality against project guidelines, provide precise written feedback, and ensure contributors follow expected quality standards.
You will assess work for statistical accuracy, data reasoning, model-selection quality, code correctness, reproducibility, metric interpretation, business-context awareness, clarity, formatting, instruction-following, and adherence to project-specific rubrics. You will spot recurring quality issues, communicate updates to trainers and QAs, support onboarding, maintain documentation, and help activate contributors who are not working consistently.
This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world’s largest AI companies and foundation-model labs. Your data science quality leadership will help ensure training data is analytically sound, reproducible, clearly explained, and aligned with client expectations.
Selection process involves an AI interview, a domain-specific task, and an interview with a recruiter.
Important: There is no immediate project for this role; however, if qualified, you will be among the first experts we reach out to when relevant opportunities arise. This will also provide you with access to future projects available through our expert network.
Requirements: - Bachelor’s, Master’s, or PhD degree in Data Science, Statistics, Computer Science, Machine Learning, Mathematics, Economics, Engineering, or a closely related quantitative field. - Strong grasp of English to follow guidelines, communicate with teams, and provide clear technical feedback. - 3+ years of professional experience in data science, analytics, machine learning, statistical modeling, experimentation, data engineering, technical review, or data science education. - Strong understanding of statistics, probability, data cleaning, exploratory data analysis, feature engineering, supervised/unsupervised learning, model evaluation, experimentation, regression, classification, clustering, and validation methods. - Ability to evaluate data science content against detailed rubrics and identify issues such as data leakage, flawed assumptions, incorrect metrics, weak methodology, non-reproducible code, hallucinated libraries/APIs, or misleading conclusions. - Familiarity with tools such as Python, pandas, NumPy, scikit-learn, SQL, Jupyter, matplotlib, R, Spark, Git, MLflow, notebooks, dashboards, and cloud/data platforms is preferred. - Experience leading or supporting remote teams of trainers, annotators, analysts, data scientists, engineers, educators, or QAs is strongly preferred. - Comfortable using Discord, Google Sheets, Google Docs, trackers, dashboards, GitHub, and project management systems. - Highly organized and able to maintain style guides, trackers, FAQs, onboarding materials, honeypots, calibration tasks, and quality documentation. - Experience with AI training, data annotation, LLM evaluation, data science QA, or rubric-based technical review is a strong plus.