Generalist
At Handshake AI, I worked on AI training and response evaluation projects focused on improving model accuracy and reasoning quality. My responsibilities included comparing multiple AI-generated responses in areas such as mathematics and image-based tasks, assessing correctness, logical reasoning, clarity, and overall response quality. I evaluated outputs using detailed rating guidelines and provided consistent feedback to support model training and optimization. This role required strong analytical thinking, attention to detail, and the ability to apply objective evaluation standards across large volumes of data. I developed experience in identifying inaccuracies, rating response quality, and ensuring consistency in AI-generated content, while working independently in a fast-paced remote environment.