AI Trainer · Data Annotator & Evaluator · Multilingual QA Specialist
Send a job offer directly to this candidate
Freelance AI evaluator with 3+ years verifying model and agent outputs against defined quality standards, including conversational agent behavior, multi-step task correctness, transcription, translation, and TTS data, across Appen/CrowdGen, Mindrift, and Toloka. Experienced writing and applying structured rubrics to judge task success, calibrating consistently across large volumes of ambiguous cases, and documenting reasoning clearly enough for other raters to follow. Certified AI Language Evaluator & Annotator (micro1).
Urdu and Punjabi speaker with C1 English, adding a linguistic and cultural accuracy layer to multilingual AI training and evaluation data.
LLM Hazard Response Evaluator - Toloka - Freelance
(2026-01)
AI Agent Evaluator - Mindrift - Freelance
(2026-07)
AI Language & Audio Evaluator - Appen/CrowdGen - Part-time
(2026-06)
BS Zoology - Zoology - University of Sargodha, Pakistan (2015-01 - 2019-01)