AI Training & Data Specialist | RLHF & Model Alignment
Send a job offer directly to this candidate
Detail-oriented AI Training Specialist and Data Evaluator with hands-on expertise in model alignment, Reinforcement Learning from Human Feedback (RLHF), and high-precision multimodal data annotation. Skilled in evaluating complex model outputs, refining prompt responses, enforcing strict taxonomic rubrics, and identifying subtle edge cases across text, video, and kinetic datasets. Proven track record of maintaining high accuracy (>98%), low error rates, and delivering high-quality training signals for frontier AI models under flexible remote workflows.
AI Evaluation & Model Alignment Specialist (2024-01 – Present)
AI Trainer & Data Annotation Specialist at Outlier AI (2024-01 – 2025-12)
Diploma in Computer Science & IT – Mount Kenya University (2023-08)