CronJobs

fullstack jobs

Staff+ Fullstack Software Engineer, Safeguards Engineering

Anthropic · San Francisco, CA | New York City, NY

hybridstaff$405,000–$405,000Posted Oct 9, 2026TypeScriptReactPython

Apply on the employer site

About this role

**About Anthropic** Anthropic’s mission is to create reliable, interpretable, and steerable AI systems—safe and beneficial for users and society. > **Note:** This job post is for **Fullstack Software Engineers** only. If you’re a **Backend Engineer**, please see the Backend Software Engineer posting. --- **About the role** Join the **Safeguards** team to help build safety and oversight mechanisms for Anthropic’s AI systems. You’ll work to **monitor models, prevent misuse, and ensure user well-being** by building systems that detect unwanted model behaviors and prevent disallowed use. You’ll uphold principles of **safety, transparency, and oversight**, while enforcing **terms of service** and **acceptable use policies**. --- **Responsibilities** - Design and build internal review tools analysts use to investigate abuse (dashboards, case queues, evidence viewers) for fast, accurate, high-volume decisions. - Build interfaces where humans supervise agents (streaming output, interruption + approval flows, and clear provenance for every agent action). - Surface abuse patterns to research teams via visualizations and exploration tools. - Own frontend architecture for tools handling sensitive data at scale (performance on large datasets, correctness with real-time updates, and end-to-end access controls). --- **You may be a good fit if you have** - A Bachelor’s degree in Computer Science / Software Engineering (or comparable experience) - Proficiency in **TypeScript** and modern **React**; comfortable in **Python** for backend work - Ability to work across the stack - Strong communication skills and ability to explain complex technical concepts to non-technical stakeholders --- **Strong candidates may also have** - 8+ years of software engineering experience (excluding internships/co-ops) - Experience building data-dense internal tools or operator interfaces (not just marketing/consumer) - Experience with integrity, spam, fraud, or abuse detection/mitigation - Experience building trust & safety detection mechanisms and intervention for AI/ML systems - Experience with prompt engineering, jailbreak attacks, and other adversarial inputs - Experience partnering closely with operational teams to build custom internal tooling --- **Compensation (Annual Salary)** - **$405,000 — $485,000 USD** --- **Logistics** - **Minimum education:** Bachelor’s degree or equivalent combination of education/training/experience - **Required field of study:** Relevant to the role (via coursework, training, or professional experience) - **Minimum years of experience:** Correlates with internal job level - **Location-based hybrid policy:** Expect to be in an office at least **25% of the time** (some roles may require more) - **Visa sponsorship:** Yes—Anthropic makes reasonable efforts to sponsor visas when possible (with support from an immigration lawyer) --- **How we’re different** Anthropic focuses on high-impact AI research as “big science,” working as a cohesive team on a few large-scale efforts. Communication and collaboration are highly valued. --- **Come work with us!** Competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a collaborative office environment. **Candidate AI Usage:** Learn about Anthropic’s policy for using AI in the application process.

Listing freshness

CronJobs last confirmed this listing 2h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord