Staff+ Software Engineer, Safeguards

Anthropic · San Francisco, CA | New York City, NY · $320k–$485k

Posted
2 days ago
Last confirmed live
Today
Published range
$320k–$485k

What this role involves

Anthropic is hiring a Staff Software Engineer for the Safeguards team to develop monitoring systems and abuse detection mechanisms for AI systems. The role requires proficiency in Python and TypeScript, and involves building automated enforcement and internal dashboards for safety oversight. Preferred experience includes trust and safety detection for AI/ML, prompt engineering, and adversarial input handling.

Skills this posting asks for

  • python
  • typescript
  • abuse detection
  • trust and safety
  • machine learning
  • prompt engineering
  • adversarial inputs
  • internal tooling

Requirements

  • Level: staff
  • Remote policy: hybrid
  • Visa sponsorship: yes

From the employer’s posting

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, en…

Read the full description on Anthropic’s careers page

Apply without filling the form

Approve this role and the application is completed for you, including a résumé tailored to it. You get a confirmation when it lands, and a credit is only spent when a submission is confirmed.

Other roles at Anthropic

All 142 roles at Anthropic