About the Role
As a Safeguards Enforcement Operations Specialist, you will be at the forefront of maintaining the integrity of our AI systems by labeling data, reviewing user interactions, enforcing policy guidelines, and collaborating across teams to continuously improve our safety mechanisms. In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a sexual, violent, or psychologically disturbing nature.
Responsibilities
- Review customer signals, user exchanges, and code to identify potential policy violations
- Enforce established Safeguards workflows and policy guidelines with precision and care
- Collaborate with Strategy Analysts, Policy Managers, and Threat Intelligence Investigators to identify emerging trends, highlight potential policy gaps, and flag content for classifier refinement
- Produce comprehensive weekly reports on operational workflows
- Handle customer communications and appeals related to policy enforcement
- Provide surge review support and content labeling when team needs arise
- Triage and respond to cross-functional requests, offering timely assistance and insights
Requirements
- A quick learner with the ability to adapt rapidly to changing environments
- Demonstrate proactivity and strong independent work capabilities
- Consistently meet deadlines and deliver high-quality work
- Possess exceptional verbal and written communication skills
- Thrive in collaborative team settings
- Have a keen eye for detail and systematic approach to complex tasks
- Show genuine interest in AI safety and ethical technology development
- Experience in content moderation, Trust & Safety, Safeguards, or related fields
- Familiarity with policy enforcement in digital platforms
- Understanding of AI ethics and responsible technology deployment
- Background in analyzing user behavior and identifying potential risks
Skills
- Content moderation
- Trust & Safety
- Safeguards
- Policy enforcement
- AI ethics
- Responsible technology deployment
- User behavior analysis
- Risk identification
- Data labeling
- AI safety
Location
- San Francisco
Work Type
- Full-time
Experience Level
- Entry level
- Mid-level
About the Company
- Employer.com is part of a family of incredible brands alongside Flawless Recruit and Recruiter.com. Together, we provide talent acquisition services to fit the unique hiring challenges of our clients; whether building recruiting processes, attracting top talent, or payrolling contractors, we can help.
- Anthropic is a public benefit corporation headquartered in San Francisco, dedicated to ensuring that artificial intelligence systems are safe and beneficial to humanity. We are a team of researchers, engineers, and policy experts working together to develop responsible AI technologies.
Equal Opportunity
- All your information will be kept confidential according to EEO guidelines.
- We encourage applications from candidates of all backgrounds. Research shows that people from underrepresented groups are more likely to experience imposter syndrome. We want to be clear: if this role excites you, we encourage you to apply, even if you don't meet every single qualification.
- Anthropic is dedicated to building a diverse, inclusive, and equitable environment that represents a variety of perspectives and experiences. We believe diverse teams create more robust, thoughtful, and innovative solutions.
