AI Products Evaluator (Freelance)
As an AI Products Evaluator, I evaluated and annotated responses generated by large language models. I provided detailed rationales for AI-generated product outputs, including prompt suggestions, text-to-speech results, and transcriptions. My work focused on ensuring the quality and appropriateness of AI responses for various use cases. • Performed evaluation of LLM responses on text-based tasks • Annotated prompts, text-to-speech, and transcription outputs • Provided rationale and quality feedback on model results • Focused on usability, context relevance, and user experience