Founder & AI Safety Researcher - Independent
Founded an AI safety and security evaluation practice focused on identifying jailbreak vectors, model misuse patterns, and adversarial prompt behavior using established testing frameworks. Designed and operated a self-hosted security monitoring stack to evaluate detection coverage against simulated attack scenarios. Applied rubric-based evaluations to assess model outputs for safety, policy compliance, and harmful-content risk while also researching defensive proxy and bot detection methods. • Evaluate LLM behavior for jailbreak and adversarial prompt risks using PyRIT, garak, and JailbreakBench • Build and run security monitoring tooling (Wazuh, Suricata, CrowdSec, Zeek, MISP) • Produce structured rationales aligned to safety and policy guidelines for consistent decisions • Manage and secure multi-server infrastructure (Proxmox, Docker, networking) supporting evaluation environments