AI Trainer / Data Annotator at Outlier (Remote)
Evaluated and rated AI-generated responses for quality, accuracy, and coherence to support LLM improvement. Applied RLHF feedback across creative writing, factual Q&A, and code explanation tasks, ensuring consistent alignment with desired outputs. Performed text classification, comparative ranking, and annotations to train and align AI models. • Rated response quality using rubric-based judgment (quality, accuracy, coherence) • Provided RLHF-style preference/feedback signals for different task types • Labeled and annotated texts for classification and comparative ranking • Maintained independent throughput and accuracy in a fully remote workflow