Outlier AI — LLM Trainer (Remote)
Trained Large Language Models by generating complex prompt sets to elicit and identify model errors. Provided concise feedback and corrected solutions to improve the model’s reasoning and response generation. Delivered results by completing many training tasks across diverse projects. • Created prompts aimed at catching and correcting specific failure modes • Iteratively refined model outputs using short targeted corrections • Evaluated response quality during training task completion • Improved reasoning quality through structured feedback cycles