AI Safety & Prompt Evaluation Specialist, Freelance (Remote)
Crafted adversarial and scenario-based prompts to evaluate LLM reasoning quality, bias, and adherence to safety constraints. Simulated dialogue scenarios to test content moderation behaviors and ethical boundary responses. Analyzed model outputs for linguistic accuracy, emotional tone, and safety/compliance violations using structured evaluation templates. • Designed prompt scenarios specifically to elicit unsafe, biased, or non-compliant responses. • Evaluated outputs with QA scoring rubrics covering language and compliance signals. • Developed creative dialogue simulations to test moderation and ethical limits. • Partnered with research teams to refine evaluation templates and improve scoring consistency.