← jobs

Cyber Evaluations Engineer

Anthropic · San Francisco, CA | Washington, DC

Builds and runs evaluations for cyber-relevant model capabilities and safeguard robustness, including jailbreak and prompt-bypass testing.

AI securityAI evaluationssafeguards

Independent aggregation of a publicly posted role. AI Compliance Index is not affiliated with, endorsed by, or recruiting for Anthropic. All applications happen on the employer's own site; listings may expire.