AI Trainer / AI Evaluator - Remote Contract (Outlier)
Evaluated AI-generated responses for accuracy, relevance, and coherence using defined quality assurance standards. Performed response evaluation to identify errors, biases, and inconsistencies and recorded findings for downstream engineering review. Designed and applied structured prompt sets to assess model reasoning and performance under varied conditions. • Annotate and label large natural language processing datasets supporting supervised and reinforcement learning pipelines • Conduct quality assurance checks aligned to evaluation metrics • Document error/bias/inconsistency patterns with clear technical notes • Deliver actionable feedback reports via technical writing to improve model outputs