Subject Matter Expert, Adversarial Red Teaming (
Jul 2024 - Current (2 years 1 month)
Designed adversarial test suites exposing harmful, biased, or deceptive model outputs.
Engineered psychologically informed jailbreaks, manipulation prompts, and obfuscation attacks.
Performed reasoning, compliance, hallucination, and privacy-risk evaluations.
Built behavioral threat models and safety rubrics used by engineering and policy teams.
Advised on mitigations to reduce systemic risk and strengthen model alignment.