Staff+ Software Engineer, Safeguards Human Review Tooling
San Francisco, California, United States · Hybrid
$320k–$485k/yr
Senior+Visa sponsorship$129B raised
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of co…
Skills: Full-stack Engineering, Platform Engineering, System Architecture, Internal Tooling, API Development
City Hall Building Manager (0923) - Real Estate Division, Office of City Administrator
San Francisco, California, United States · On-site
Mid level
Company Description Permanent Exempt (PEX): Full Time position is excluded by the Charter from the competitive civil service examination process and shall serve at the discretion of the appointment officer. The anticipat…
Hotel Maintenance Manager | Axiom Hotel | San Francisco, CA | Modus by PM Hotel Group
San Francisco, California, United States · On-site
$85k–$95k/yr
Mid level
Axiom Hotel in San Francisco is looking for skilled, detail oriented and hardworking Maintenance Manager to join our team! Pay Range - $85,000 - $95,000 annually ABOUT AXIOM HOTEL: Located in Union Square, within steps o…
Skills: Preventive maintenance, Vendor management, Team leadership, Budgeting, Staff training
San Francisco, California, United States · On-site
$134k–$162k/yr
Mid level
Company Description The Department of Public Health prioritizes equitable and inclusive access to quality healthcare for its community and values the importance of diversity in its workforce. All employees at the Departm…
San Francisco, California, United States · On-site
$131k–$159k/yr
Mid level
Company Description The Department of Public Health prioritizes equitable and inclusive access to quality healthcare for its community and values the importance of diversity in its workforce. All employees at the Departm…
San Francisco, California, United States · On-site
$152k–$178k/yr
Senior$2.0B raised
Come join the organization that is redefining security for the AI era. As one of the fastest-growing startups ever, we enable teams to secure cloud and AI applications by connecting code, cloud, and runtime into a single…
Senior Customer Engineer, Enterprise AI - Bay Area
San Francisco, California, United States · Hybrid
$231k–$300k/yr
Senior$4.3B raised
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging…
Skills: Web Security, Networking, Cloud Infrastructure, AI-Native Development, Technical Discovery
Principal Account Executive, SLED West (R1 Institutions and Higher Ed)
San Francisco, California, United States · On-site
Mid level$4.3B raised
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging…
Senior Account Executive, Startups (San Francisco)
San Francisco, California, United States · Hybrid
$244k–$336k/yr
Senior$4.3B raised
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging…
About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging…
San Francisco, California, United States · On-site
$206k–$283k/yr
Senior$32B raised
GAQ327R194 As an industry leader in data and AI, Databricks transforms how global enterprises operate. To support our rapid growth, we are hiring a Workplace Services Manager to drive workplace operations at our San Fran…
Skills: Workplace operations, Team leadership, Vendor management, Budgeting, Strategic planning
San Francisco, California, United States · On-site
$202k–$236k/yr
Senior$665M raised
About Us Beast Industries is a multifaceted media and entertainment company founded by Jimmy Donaldson, popularly known as MrBeast, the most watched person in the world. Renowned for revolutionizing digital content creat…
Skills: Game security, Anti-cheat engineering, Python, TypeScript, API security
At Early Warning, we’ve powered and protected the U.S. financial system for over thirty years with cutting-edge solutions like Zelle®, Paze℠, and so much more. As a trusted name in payments, we partner with thousands of …
Director of Enterprise Safety & Strategic Operations
San Francisco, California, United States · On-site
$180k–$185k/yr
Senior+
Note:If you are a current employee, please apply through the internal career page. IOA is on the forefront of revolutionary healthcare models, reshaping the way people can age in place. Our innovative models transform li…
About Sela Sela is the market-leading AI voice agent for mortgage lenders. Founded in 2024, we’ve grown like crazy (~$0->13M in 20 months), and currently count 5 of the 10 largest independent mortgage banks in the U.S. a…
Ironclad is the leading AI contracting platform that transforms agreements into assets. Contracts move faster, insights surface instantly, and agents push work forward, all with you in control. Whether you’re buying or s…
About Corridor Software development is becoming autonomous. The fastest teams already work this way, and at Corridor, we are leading the way on secure development as the rest of the industry follows. Security tools were …
Skills: Corporate IT, IT Security, Systems Administration, Identity And Access Management, Endpoint Security
San Francisco, California, United States · On-site
$21/hr–$23/hr
Entry level
Note:If you are a current employee, please apply through the internal career page. IOA is on the forefront of revolutionary healthcare models, reshaping the way people can age in place. Our innovative models transform li…
San Francisco, California, United States · On-site
$347k–$445k/yr
Senior$201B raised
We are hiring a Security Software Engineer to design and implement the hardware-backed security foundations used across OpenAI’s device ecosystem. A central focus of this role is hardening the boundary between our policy…
San Francisco, California, United States · On-site
$23/hr
Entry level
Job Description: Job Description: Office Agents are responsible for helping customers during the shipping process by coordinating shipping details, completing compliance documentation, and providing customer service. All…
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
$320k–$485k/yr
Full-time
bachelor degree
Competitive Compensation, Optional Equity Donation Matching, Generous Vacation, Parental Leave, Flexible Working Hours, Office Space
Visa sponsorship available
Posted 7d ago
~40 hrs/week
Responsibilities
Build and scale investigation, review, and enforcement tooling to identify harmful behavior across first and third-party platforms. Develop the underlying platform layer of APIs and services while integrating LLMs to automate review workflows.
Requirements
Requires a technical background in full-stack or platform engineering with experience shipping internal tools for operational users. Preferred candidates have 8+ years of experience and a background in trust and safety or abuse-prevention systems.
Full job description
About Anthropic
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
About the role
The Safeguards team is responsible for ensuring our models and products are developed and deployed safely. We're looking for engineers for our Review Tooling team, which builds the systems that humans — and increasingly Claude — use to investigate potential harms and take enforcement actions across Anthropic's first-party products and third-party cloud platforms.
This is a foundational role: as one of the first engineers on this new team, you'll own the tools our safety investigators rely on to understand what's happening on our platforms and act on it, as well as the platform underneath those tools. That platform includes analytics capabilities, privacy-preserving primitives that keep review workflows compatible with our data retention commitments, and a sandbox environment where new review interfaces and workflows can be built and iterated quickly. As model capabilities and usage grow, you'll also drive how we scale review through automation — building systems where Claude meaningfully extends what human reviewers can do, while keeping people in the loop where their judgment matters most.
These are internal tools, but they are anything but low-stakes: the speed, clarity, and reliability of this tooling directly determines how quickly Anthropic can identify harmful behavior, make sound enforcement decisions, and feed signal back into model training. You'll partner closely with policy, operations, data science, legal, and privacy teams to ensure our enforcement systems are effective, accurate, and trustworthy.
Key responsibilities
Build investigation, review, and enforcement tooling for both first-party and third-party platform surfaces — including case queues, investigation views, decision and audit logging, and account-actioning workflows
Develop the platform layer of reusable APIs, data storage, and backend services that lets new review workflows be stood up quickly and safely
Scale review through automation, including enabling reviewers to use Claude effectively and building toward Claude-assisted and Claude-driven review workflows
Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable, well-designed systems that measurably reduce handling time and decision error
Build in the guardrails that sensitive internal tools require: granular permissions, audit trails, data-access controls, and reviewer wellbeing features such as content obfuscation and exposure limits
Instrument the tools you ship — surfacing metrics on queue health, reviewer throughput, and decision quality — and ensure tooling evolves alongside new privacy primitives and data retention commitments
Minimum qualifications
A technical background in full-stack or platform engineering, with the ability to engage deeply in architecture and design discussions
Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably
Experience working cross-functionally with non-engineering partners such as operations, policy, or legal teams
Excellent communication skills, including the ability to explain technical tradeoffs to non-technical stakeholders
Care about the societal impacts of AI and want your work to make powerful systems safer
Preferred qualifications
8+ years of industry software engineering experience
Experience building trust and safety, integrity, fraud, or abuse-prevention tooling, or other systems supporting human review at scale
Experience designing systems under strict privacy, compliance, or data governance constraints, such as zero data retention environments
Experience integrating LLMs or agentic systems into operational workflows, or building human-in-the-loop automation — including using agentic coding tools (e.g., Claude Code) as a core part of your own development workflow
Experience building developer platforms or extensible tooling frameworks that other teams build on top of
Experience supporting enforcement or moderation systems across multiple product surfaces, including enterprise or cloud platform contexts
A product-minded approach to internal users: you work directly with the people using your tools, watch where they struggle, and fix it
The annual compensation range for this role is listed below.
For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role.
Annual Salary:
$320,000—$485,000 USD
Logistics
Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices.
Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team.
Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings.
How we're different
We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills.
The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences.
Come work with us!
Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
Related keywords
Software EngineeringSafeguardsTrust and SafetyLLMClaudeFull-stackPlatform EngineeringAPIData RetentionPrivacyComplianceAutomationAgentic SystemsAudit TrailsContent ObfuscationEnforcement Systems
Anthropic is an AI safety and research company working to build reliable, interpretable, and steerable AI systems.
Industry
Research Services
Company size
501-1,000 employees
LinkedIn followers
4,444,172
Total funding
$129B
We're an AI research company that builds reliable, interpretable, and steerable AI systems. Our first product is Claude, an AI assistant for tasks at any scale.
Our research interests span multiple areas including natural language, human feedback, scaling laws, reinforcement learning, code generation, and interpretability.
GamingEducationSoftwarePredictive AnalyticsInternet of ThingsOnline GamesManufacturingIncubatorsInformation TechnologyMedical Device
How much do Security & Safety jobs in San Francisco, CA pay?
Based on 1002 listings with disclosed salaries, most security & safety jobs in San Francisco, CA pay between $75k–$270k per year. Individual offers vary with seniority, company size, and specialization.
How many Security & Safety jobs are open in San Francisco, CA right now?
There are currently 1,461 open security & safety positions in San Francisco, CA listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Security & Safety roles in San Francisco, CA?
Companies currently hiring include Allied Universal, OpenAI, San Francisco Department of Public Health, Anthropic, UCSF Department of Anesthesia and Perioperative Care, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Security & Safety jobs in San Francisco, CA?
Yes — 525 of the 1461 open security & safety positions offer remote or hybrid work (100 remote, 425 hybrid).
How do I apply for Security & Safety jobs in San Francisco, CA?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.