← jobs
Cyber Evaluations Engineer
Anthropic · San Francisco, CA | Washington, DC
Builds and runs evaluations for cyber-relevant model capabilities and safeguard robustness, including jailbreak and prompt-bypass testing.
AI securityAI evaluationssafeguards
Independent aggregation of a publicly posted role. AI Compliance Index is not affiliated with, endorsed by, or recruiting for Anthropic. All applications happen on the employer's own site; listings may expire.