AI Safety Researcher: Safety, Evaluation & Red-Teaming

Posted 2 days ago

thinking machines labMillbrae (CA)

SENIORITY

Junior

Apply

About the role

Thinking Machines Lab Inc. in San Francisco is seeking a safety researcher to bridge research and hands-on engineering, focusing on making models safe and trustworthy. You will explore how training shapes refusals and how to evaluate boundaries, designing experiments to inform model training and evaluation strategies. Join a team that values rigorous analysis, data-driven evaluation, and responsible AI practices, with opportunities to influence real-world deployments.

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this