Skip to main content

36 AI Safety jobs in United States

Showing 1–20 of 36 jobs

H

Product Policy Manager, Product Risk

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$245–285K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

2 mo. ago

H

Applied AI Architect, International Policy

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$240–315K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

1 wk. ago

H

Research Manager, Biological Safety

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$405–485K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

1 mo. ago

H

Head of Policy Design, Societal Harms

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$330–395K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

1 mo. ago

H

Engagement Manager, Applied AI

https://www.anthropic.com/careers/jobs

Austin, Texas, US (Hybrid)

$275–380K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

9 mo. ago

H

Product Manager, Safeguards (Generalist)

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$385–460K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

3 wk. ago

H

Product Manager, Safeguards Rare Harms

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$305–385K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

4 mo. ago

H

Research Scientist, Interpretability

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$350–850K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

9 mo. ago

H

Staff+ Software Engineer, Safeguards Data

https://www.anthropic.com/careers/jobs

New York, US (Hybrid)

$320–485K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

2 mo. ago

H

Safeguards Enforcement Lead, User Well-Being

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$285–330K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

1 mo. ago

H

Communications Manager, Safeguards

https://www.anthropic.com/careers/jobs

San Francisco, California, US (On-site)

$265–295K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

4 days ago

H

Safeguards Enforcement Lead, Cyber Harms

https://www.anthropic.com/careers/jobs

San Francisco, California, US (Hybrid)

$285–330K / year

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.

1 mo. ago

Related searches