Ekemini Mendie
RLHF Specialist & Senior Software Engineer
I specialise in Reinforcement Learning from Human Feedback (RLHF) and AI alignment, evaluating, ranking and annotating large language model outputs to make models safer, more helpful and better at following instructions. I bring 10+ years of software engineering across fintech, security and AI-powered products, so I can review complex code generation, multi-turn dialogue and reasoning tasks in depth.
