Technical Writer & AI Training Specialist (AI Red-Teaming; SFT/RLHF/DPO-alignment workflows)
Conducted AI safety and LLM alignment work by developing and refining prompt/red-team scenarios for cybersecurity risks like prompt injection and jailbreaking. Evaluated text-based model outputs for accuracy and safety and authored rationales for ranking decisions and correction of hallucinations. Performed pairwise ranking, response auditing, and edge-case labeling to strengthen safety guardrails. • Developed technical cybersecurity/AI training content, including injection and jailbreaking cases • Ran independent red-teaming labs focused on prompt injection and jailbreaking techniques • Audited and improved model responses for accuracy and safety • Labeled edge cases and supported dataset quality via pairwise ranking