AI Trainer and Data Annotator (LLM optimization, prompt/response writing, and dataset annotation)
Delivered RLHF-based evaluation and refinement of Large Language Model outputs to enhance accuracy, safety, and conversational quality. Authored target responses and crafted complex prompts to improve AI reasoning and natural language processing performance. Audited model outputs against alignment and compliance guidelines to ensure data integrity and training readiness. • Evaluated LLM responses and iterated to maximize quality • Designed prompts and produced high-quality target responses • Tagged, categorized, and annotated large-scale training datasets • Performed compliance and alignment checks on model outputs