Handshake AI
Conducted Reinforcement Learning from Human Feedback (RLHF) to evaluate and optimize multi-modal AI model outputs based on visual data inputs. Analyzed generated image captions, structural descriptions, and spatial reasoning tasks against strict accuracy and formatting guidelines to train the model on visual comprehension.