AI Red Team Engineer - Adversarial Security (Remote)

Posted yesterday

aicomplianceMillbrae (CA)

SENIORITY

Senior

SALARY

$320,000 per year

Apply

About the role

Anthropic in San Francisco, CA is seeking a Safeguards Engineer to conduct adversarial testing across deployed AI products and capabilities. The role is remote-friendly with travel required, offering a compensation range of $320,000 to $405,000 per year. The successful candidate will identify vulnerabilities, craft attack scenarios, and collaborate with product and safety teams to harden systems while maintaining rigorous compliance with AI safety practices.

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this