AI Training Annotator (Aether Project)
Specialized in the alignment and fine-tuning of frontier Large Language Models (LLMs) via Reinforcement Learning from Human Feedback. Evaluated and audited multi-turn dialogues, ensuring context and logical consistency in model responses. Conducted adversarial and instruction-following evaluations across multimodal datasets. • Designed test prompts to identify model hallucinations. • Performed high-level evaluation of model outputs. • Executed audio-to-text and video-to-text assessment tasks. • Assessed cross-modal factual consistency in outputs.