Cross-domain AI evaluation expert (alignment, prompt engineering review, adversarial testing, red-teaming validation)
Led cross-domain evaluation activities emphasizing AI alignment, model safety, and adversarial robustness testing through systematic assessment of model outputs. • Defined evaluation criteria spanning technical accuracy, logical coherence, and human-level cognitive alignment • Performed red-teaming-style validation to surface failure modes such as jailbreak susceptibility • Supported prompt engineering review to improve reliability and reduce unsafe behaviors • Assessed responses for alignment with scientific/engineering/philosophical and behavioral expectations.