About the Role
This role is at the forefront of maintaining AI system integrity by labeling data, reviewing user interactions, and enforcing policy guidelines. You will collaborate across teams to improve safety mechanisms. Please note, this role may involve exposure to explicit content.
Responsibilities
- Review customer signals, user exchanges, and code to identify potential policy violations
- Enforce established Safeguards workflows and policy guidelines with precision and care
- Collaborate with Strategy Analysts, Policy Managers, and Threat Intelligence Investigators to identify emerging trends, highlight potential policy gaps, and flag content for classifier refinement
- Produce comprehensive weekly reports on operational workflows
- Handle customer communications and appeals related to policy enforcement
- Provide surge review support and content labeling when team needs arise
- Triage and respond to cross-functional requests, offering timely assistance and insights
Requirements
- A quick learner with the ability to adapt rapidly to changing environments
- Demonstrate proactivity and strong independent work capabilities
- Consistently meet deadlines and deliver high-quality work
- Possess exceptional verbal and written communication skills
- Thrive in collaborative team settings
- Have a keen eye for detail and systematic approach to complex tasks
- Show genuine interest in AI safety and ethical technology development
- Experience in content moderation, Trust & Safety, Safeguards, or related fields
- Familiarity with policy enforcement in digital platforms
- Understanding of AI ethics and responsible technology deployment
- Background in analyzing user behavior and identifying potential risks
Skills
- AI safety
- Ethical technology development
- Content moderation
- Trust & Safety
- Safeguards
- Policy enforcement
- AI ethics
- Responsible technology deployment
- User behavior analysis
- Risk identification
- Data labeling
- Communication
Location
- On-site
- Remote
Work Type
- Contract
- Full-time
Experience Level
- Mid-level
Salary/Compensations
- USD 145 - USD 145 - hourly
About the Company
- Employer.com is part of a family of incredible brands alongside Flawless Recruit and Recruiter.com. Together, we provide talent acquisition services to fit the unique hiring challenges of our clients; whether building recruiting processes, attracting top talent, or payrolling contractors, we can help.
- Anthropic is a public benefit corporation headquartered in San Francisco, dedicated to ensuring that artificial intelligence systems are safe and beneficial to humanity. We are a team of researchers, engineers, and policy experts working together to develop responsible AI technologies.
Equal Opportunity
- All your information will be kept confidential according to EEO guidelines.
- We encourage applications from candidates of all backgrounds. Research shows that people from underrepresented groups are more likely to experience imposter syndrome. We want to be clear: if this role excites you, we encourage you to apply, even if you don't meet every single qualification.
- Anthropic is dedicated to building a diverse, inclusive, and equitable environment that represents a variety of perspectives and experiences. We believe diverse teams create more robust, thoughtful, and innovative solutions.
