AI Training / Model Evaluation (RLHF and Prompt Engineering)
As an AI Training professional, I review and provide feedback on outputs generated by large language models. I ensure data integrity and high-quality reinforcement learning from human feedback (RLHF) to enhance model responses. My work involves prompt engineering, fact-checking, and evaluating AI model outputs for accuracy and reliability. • Evaluated a variety of LLM-generated texts for correctness and relevance. • Applied structured feedback for model fine-tuning using RLHF methodologies. • Developed and reviewed prompts to optimize LLM engagement and accuracy. • Maintained meticulous records to uphold dataset integrity and annotation quality.