AI Trainer & Language Model Evaluator (Contractor)
Served as an AI trainer and language model evaluator, using RLHF workflows to score and rank LLM response sequences. Built supervised fine-tuning (SFT) ground-truth interaction datasets by authoring error-free text pairings across technical and conversational verticals. Performed adversarial prompt simulation to identify, document, and patch alignment blind spots, bias, and factual hallucinations. • Evaluated factual precision, logical constraints, and tone for response ranking • Produced SFT training data text pairs for dense domain coverage • Ran targeted adversarial prompt simulation matrices for red-teaming alignment • Updated datasets to mitigate hallucinations and improve model alignment quality