
Explore active software engineering, data, AI, product, and design roles directly from verified employers—and make sure your resume is ready before applying.
Get an instant ATS score, missing keyword alert, and bullet rewrites tailored to your target job before submitting your application.
About the Team The Safety Training research team aims to fundamentally advance our capabilities for precisely implementing safe behavior in AI models, and to leverage these advances to make OpenAI’s deployed models safe and beneficial. This requires a breadth of new ML research to address the growing set of safety challenges as AI becomes more powerful and used in more settings. Key focus areas include how to train nuanced safety behaviors, how to make the model robust to bad actors, how to address privacy and security risks, and how to make the model trustworthy in safety-critical situations. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. About the Role We’re seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You’ll advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities. In this role, you will: Research and implement methods for safety training, reinforcement learning, and adversarial robustness. Develop evaluations, identify model failure modes, and use findings to improve training. Work with research, engineering, security, and policy partners to support safe, reliable deployment. You might thrive in this role if you: Bring 4+ years of relevant AI safety research experience, including RLHF, adversarial training, or robustness. Have a degree in computer science, machine learning, or a related field, and strong deep learning research or engineering skills. Have experience improving model safety for deployment and enjoy collaborative research. Are motivated by OpenAI’s mission and the responsible use of AI in safety-critical settings. Security Requirements Active TS/SCI clearance or equivalent. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About the Team The Cooperative AI team is scaling OpenAI with OpenAI. We are building a model-powered scaled automated workforce and knowledge system that evolves and learns alongside a human workforce. By leveraging OpenAI’s state-of-the-art models and technologies, some already in production, others still in the lab, we develop systems that reason and work autonomously for a wide variety of operational work. We leverage real workloads for critical systems across finance, sales, customer support, integrity, product insights, internal operations, and more in order to drive insights into product and industry. We partner closely with internal teams and external customers globally, operating in a hyper-fast feedback loop where many of our users are just a few steps away. This proximity allows us to iterate quickly, validate impact in real time, and accelerate industry impacting learnings and systems builds. We are a highly multidisciplinary, self-contained team focused on transforming the workplace via smart systems, knowledge, scalable and reliable primitives that apply world-class AI capabilities across domains. Our mission is to learn fast and transform how humans collaborate with AI at scale. About the Role We are looking for a hands-on Engineering Manager to lead a small, fast-moving team building AI-powered automation systems that redefine how work gets done across OpenAI. This role sits at the intersection of applied AI, research, and product engineering. You’ll lead a team that builds systems that know how to learn from humans, and carry real workloads across, sales, support, finance, IT, and more, while staying deeply involved in the technical work. You will operate in a highly iterative environment, deploying systems directly to internal users, gathering rapid feedback, and evolving solutions in real time. This is a high-ownership role for someone excited about building 0→1 systems, working closely with customers, and shaping how AI transforms operational work at scale. What You’ll Do Lead and grow a small team building applied AI systems for internal operations Design and build AI-powered automation systems in close proximity to customers Stay hands-on in architecture and implementation across the full stack Develop evolving systems spanning developer tools, automation platforms, knowledge graphs, and data systems Deploy systems directly to internal users and close customers to iterate rapidly based on real-world feedback Engage frequently with scaled workforces to understand needs and validate solutions Create systems for visibility and learning in hybrid workforces Partner with product, research, and ops teams daily You Might Thrive in This Role If You Have 12+ years of experience in engineering, including 3+ years of experience in engineering management, and at least 7 years as an IC engineer Are a hands-on builder who enjoys operating across the stack Have deep experience applying AI, and are ready to experiment with frontier approaches Bring strong technical judgment across systems design, infrastructure, and full-stack development Have built developer tools, internal platforms, or workflow automation systems Enjoy frequent interaction with customers and thrive in feedback-driven development loops Are comfortable with ambiguity and excited to operate in a rapidly evolving space Have experience in operationally complex environments (e.g., logistics, support systems, internal tooling, warehouse automation, etc.) About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer, and we do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. For additional information, please see OpenAI’s Affirmative Action and Equal Employment Opportunity Policy Statement . Background checks for applicants will be administered in accordance with applicable law, and qualified applicants with arrest or conviction records will be considered for employment consistent with those laws, including the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance for Employers, and the California Fair Chance Act, for US-based candidates. For unincorporated Los Angeles County workers: we reasonably believe that criminal history may have a direct, adverse and negative relationship with the following job duties, potentially resulting in the withdrawal of a conditional offer of employment: protect computer hardware entrusted to you from theft, loss or damage; return all computer hardware in your possession (including the data contained therein) upon termination of employment or end of assignment; and maintain the confidentiality of proprietary, confidential, and non-public information. In addition, job duties require access to secure and protected information technology systems and related data security obligations. To notify OpenAI that you believe this job posting is non-compliant, please submit a report through this form . No response will be provided to inquiries unrelated to job posting compliance. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made via this link . OpenAI Global Applicant Privacy Policy At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Analyst focused on Violence & Extremism, you will be responsible for building and executing operational workflows to assess model behavior, drive enforcement decisions, and develop evals across a technically demanding range of policy areas. Your work spans detecting and mitigating attempts to misuse Anthropic's AI systems to facilitate real-world harm, including weapons and dangerous technology, critical infrastructure attacks, violent extremism, and threats of violence. Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, graphic, hateful, or psychologically disturbing nature. Key responsibilities Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy Develop and maintain evals that measure model performance on these policy areas, surface regressions, and inform policy and model improvements Partner with Engineering and Data Science to optimize detection and automated enforcement systems for potential policy violations Review flagged content to drive enforcement decisions and surface policy gaps, with particular attention to novel or technically sophisticated misuse attempts + emerging extremist movements, ideologies, and mobilization tactics Support the Safeguards policy design team by providing structured feedback on policy gaps and enforcement ambiguities based on real enforcement scenarios Develop and maintain enforcement guidelines and reviewer documentation that enable accurate, consistent enforcement across a wide range of content Keep up to date with emerging threats, terrorist and extremist movements, regulatory changes, and AI policy enforcement best practices, and apply these to inform our workflows and evals Identify and escalate emerging misuse patterns, novel attack vectors, and signs of coordinated violent extremist activity Minimum qualifications Experience in policy enforcement, threat intelligence, counterterrorism, government, or a closely related field, with direct exposure to harmful content, dangerous technology, violent extremism, or physical harm facilitation Experience standing up and scaling policy enforcement or content review workflows Proficiency in SQL and/or other data analysis tools to draw insights from large datasets and monitor enforcement workflow health Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams Experience working with generative AI products, including writing effective prompts for content review and enforcement Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space Preferred qualifications Subject matter expertise in one or more high-stakes harm areas, such as weapons and dangerous technology, violent extremism, terrorism, autonomous systems, or critical infrastructure protection Familiarity with relevant legal and regulatory frameworks governing dangerous technology, critical infrastructure, or domestic/international terrorism Experience developing evals or red-teaming AI systems, particularly for harmful content or policy enforcement use cases Experience with threat actor profiling and threat intelligence frameworks (e.g., MITRE ATT&CK) Experience tracking threat actors, extremist networks, or misuse patterns across surface, deep, and dark web environments Experience with large language models and an understanding of how AI technology could provide meaningful uplift toward serious harm Proficiency in Python for data analysis and workflow automation Background in law enforcement, national security, defense, counterterrorism, or a relevant regulatory environment Experience assessing the technical plausibility and real-world harm potential of content, including the ability to distinguish between general educational content and genuine operational uplift, and between protected speech and genuine incitement/mobilization Familiarity with cross-platform threat analysis and OSINT techniques The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $285,000 — $330,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Analyst on the User Well-being team, you will be focused on supporting the design and deployment of mental health guardrails – iterating on detection systems, managing review queues, evaluating new interventions, and monitoring existing ones. Interventions range from steering how Claude responds in the conversation itself to in-product features that connect users to other resources. This work involves translating expert clinical guidance, internal data analyses, and legal or engineering constraints into concrete changes to how we detect, review, and respond. This team covers a broad set of interconnected harms including suicide, self-harm, disordered eating, AI sycophancy, and emotional dependence on AI. This position may expand into broader areas of user well-being enforcement over time. Safety is core to our mission, and you'll help shape policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way. *Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a sexual, violent, or psychologically disturbing nature. Key responsibilities Support the design and execution of interventions, defining key metrics, and curating evaluation datasets Partner with Engineering and Data Science teams to build, tune, and validate detection models for automated intervention systems, including threshold-setting and precision and recall tradeoffs Monitor how interventions and detection systems perform over time Review flagged content to drive enforcement and policy improvements Support the development of in-product features that connect users to crisis resources, working with Product, Legal, and external partners on referral pathways and user-facing content Support the Safeguards Policy Design team by providing detailed feedback on policy gaps based on real scenarios Keep up to date with emerging AI policy and external research on AI's relationship to mental health, and use these to inform our decision-making and workflows Minimum qualifications Experience in trust & safety, product policy, content moderation, or a related field, with direct exposure to mental health, suicide and self-harm, or related well-being harm areas Experience designing or running experiments, evaluations, or measurement studies to determine whether an intervention worked Experience translating policy definitions into measurable form — the rubrics, review guidelines, or classification criteria, whether applied by human reviewers or automated systems Experience managing or coordinating content review operations, including quality assurance and workflow management Proficiency in SQL and/or other data analysis tools to measure intervention efficacy, monitor workflow health, and surface policy gaps Experience working with generative AI products, including writing effective prompts for content review, classification, or evaluation Experience turning open questions and data into concise and insightful analysis Experience identifying emerging risks and communicating findings to cross-functional stakeholders Understanding of the challenges involved in implementing product policies at scale in the content moderation space Sound judgment in ambiguous, high-consequence cases, and comfort making a call and escalating appropriately when the available signal is incomplete Preferred qualifications Subject matter expertise in mental health, whether developed in academia, clinical practice, crisis intervention, trust & safety, or other related settings Experience building or evaluating LLM-based classification systems Experience using agentic tools (e.g. Claude Code) to scale analysis or automate recurring work Experience working within crisis support The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $245,000 — $285,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the Role Anthropic's Safeguards team is responsible for enforcing our policies, protecting users, and ensuring our platform is not misused. As a Safeguards Enforcement Analyst focused on Safety Evaluations, you'll play a central role in ensuring our models meet safety and policy standards before and after launch. You'll run and monitor evaluations, drive mitigations when issues surface, coordinate the creation of new evals, and help build the processes and documentation that allow the team to scale this work over time. This role requires someone who is detail-oriented, comfortable navigating ambiguity, and capable of coordinating across teams to break new ground and drive work to completion. This work is deeply cross-functional — you'll partner closely with policy experts, Safeguards engineering teams, and many other stakeholders throughout the organization to ensure our evaluations are comprehensive and current, and that findings translate into meaningful improvements to model behavior. Responsibilities Support model launch readiness by running evaluations, monitoring and interpreting results, and surfacing regressions or unexpected behavior changes to relevant stakeholders Partner closely with policy and domain experts throughout the evaluation lifecycle — from identifying risks and scoping the right evaluation approach, to coordinating creation of new evals and ensuring existing ones remain current with evolving policies, threat vectors, and model capabilities Work with cross-functional stakeholders to help manage evaluation outcomes, including interpreting results and driving mitigations where needed Think strategically about eval quality to build processes and eval paradigms that keep evaluations unsaturated, high-signal, and insightful as models improve Build out processes and frameworks for creating product-specific evaluations as Anthropic's product surface area expands Help design and scope tooling improvements that accommodate evolving eval needs and expand self-serve eval creation and iteration for non-technical users Write and maintain rigorous documentation for evaluation creation, execution, and interpretation as the team builds out eval tooling and processes You may be a good fit if you: Have experience in trust and safety, content operations, policy enforcement, or a related operational role at a technology company Thrive in ambiguous, fast-moving environments — you're energized rather than frustrated when the path forward isn't clearly defined and you need to figure it out as you go Have experience building processes, workflows, or programs from scratch (zero-to-one work), not just maintaining existing ones Have strong program management instincts, naturally creating structure around complex, multi-stakeholder efforts by tracking timelines, dependencies, and deliverables to keep work on track Are eager to expand your technical toolkit, including adopting internal tools and AI-assisted workflows (e.g., Claude Code) to accelerate your work Can manage multiple concurrent workstreams across different domain areas without losing track of details — strong prioritization and context-switching are essential when deadlines and priorities shift quickly Are a strong generalist comfortable moving fluidly across different types of work and switching contexts throughout the day Are comfortable making judgment calls with incomplete information and escalating appropriately when needed Communicate clearly and concisely, both in writing and cross-functionally Strong candidates may also have: Experience operating under tight, high-stakes timelines — such as product launch cycles, incident response, or regulatory deadlines — where information and priorities can shift with little notice Experience coordinating across engineering, policy, and product teams to translate findings into concrete action Experience building and maintaining SOPs, runbooks, and operational documentation in fast-changing environments Proficiency with data tools (SQL, dashboards, spreadsheets) sufficient to maintain and improve workflows Comfort working with sensitive content areas as part of eval creation or enforcement review responsibilities The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $230,000 — $270,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Analyst focusing on Integrity & Authenticity, you will be responsible for building and executing enforcement workflows for our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for coordinated inauthentic behavior, election manipulation, and targeting, tracking, and surveillance of individuals. Your work will span a broad and interconnected set of harm areas: AI-enabled influence operations and disinformation campaigns, the abuse of AI to interfere with electoral processes, and the use of AI systems to facilitate stalking, surveillance, profiling, and the targeting of individuals or groups. Safety is core to our mission, and you'll help shape policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way. Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a political, violent, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays, particularly around major electoral events. Key responsibilities Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy Partner with Engineering and Data Science teams to optimize detection models for policy violations and automated enforcement systems Review flagged content to drive enforcement and policy improvements Enforce usage policies with a focus on detecting and mitigating AI-enabled influence operations, coordinated inauthentic behavior, election interference, and targeting, tracking, or surveillance of individuals and groups Support the Safeguards policy design team by providing detailed feedback on policy gaps based on real enforcement scenarios Keep up to date with emerging AI policy enforcement best practices, evolving threat actor tactics, and the regulatory landscape around elections, privacy, and surveillance, using these to inform our decision-making and workflows Minimum qualifications Experience in trust & safety, policy enforcement, threat intelligence, or a closely related field with a focus on one or more of: influence operations, disinformation, coordinated inauthentic behavior, election integrity, or privacy and surveillance harms Experience standing up and scaling policy enforcement or content review workflows Proficiency in SQL and/or other data analysis tools to draw insights from large datasets Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams Experience working with generative AI products, including writing effective prompts for content review and enforcement Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space Preferred qualifications Experience conducting cross-platform investigations into influence operations, coordinated inauthentic behavior, or disinformation campaigns Familiarity with open-source intelligence (OSINT) techniques and tools used for threat actor tracking and network analysis Working knowledge of privacy law, surveillance technology, or data broker ecosystems as they relate to targeting and tracking harms Experience with large language models and an understanding of how AI technology could be misused to generate synthetic personas, fabricate quotes, or automate persuasion at scale Familiarity with election security frameworks, campaign finance law, or electoral integrity standards in one or more jurisdictions Experience navigating evolving regulatory landscapes relevant to this space (e.g., DSA, EU AI Act, FEC regulations, GDPR) Experience working with election bodies, civil society organizations, or government agencies on integrity or disinformation-related issues Proficiency in Python for data analysis and automation Experience with dark web monitoring or tracking threat actors across surface, deep, and dark web environments The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $285,000 — $330,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As an Enforcement Analyst, you will be responsible for reviewing content and executing enforcement actions across our products and services, with a focus on detecting and mitigating attempts to misuse Anthropic's AI systems for malicious cyber operations. Your initial focus will center on reviewing flagged activity related to cyberattacks, malware development, and offensive exploitation; however, this position may later expand to include broader areas of enforcement. Safety is core to our mission, and you'll help uphold policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way. Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, technical, or psychologically disturbing nature. This role may require responding to escalations during weekends and holidays. Key responsibilities Review flagged content and accounts to make accurate, well-documented enforcement decisions in line with our usage policies Detect and mitigate potential misuse of AI systems to facilitate cyberattacks, malware creation, exploitation tooling, and related harmful cyber operations Triage and escalate novel, ambiguous, or high-severity cases to appropriate stakeholders Provide detailed feedback to the Safeguards policy design team on policy gaps surfaced through real enforcement scenarios Partner with Engineering and Data Science teams by surfacing detection model errors and quality signals from review to improve precision and recall Maintain high accuracy and consistency standards across review queues Keep up to date with emerging AI policy enforcement best practices, threat actor tactics, and the evolving cyber threat landscape, using these to inform enforcement decisions Minimum Qualifications Experience in cybersecurity, including knowledge of offensive techniques, exploit development, malware analysis, or vulnerability research Experience performing content review, abuse investigations, or policy enforcement at volume Proficiency in SQL and/or Python for data analysis and threat detection Experience identifying emerging risks and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams Experience working with generative AI products, including writing effective prompts for content review and enforcement Preferred qualifications Experience in trust & safety, abuse investigations, cybersecurity investigations, or threat intelligence in a technology or AI company Experience with large language models and an understanding of how AI technology could be misused for cyber operations Experience operating within abuse monitoring programs or enforcement review systems Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space Experience working with government agencies, regulated environments, or information sharing communities The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $285,000 — $330,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Analyst on the Child Safety team, you will be responsible for our child safety enforcement workflows, responsible for scaling, maintaining, and continuously improving the systems and processes we use to detect and respond to child sexual abuse material (CSAM) and child sexual exploitation material (CSEM) generated or facilitated through Anthropic's products. This is a deeply operational role. You will serve as the central point of contact for those who conduct content review, managing day-to-day workflows, quality assurance, and escalation processes to ensure reviews are accurate, consistent, and conducted with appropriate support structures in place. You will also work closely with internal Engineering, Policy, and Legal teams to scale detection systems and close enforcement gaps as the threat landscape evolves. This work is essential to Anthropic's mission. Child safety is one of our highest-priority harm areas, and the person in this role will have a direct and meaningful impact on protecting children from AI-facilitated exploitation and abuse. Important context for this role: In this position you will regularly be exposed to and engage with explicit content of a sexual nature involving minors, as well as content that may be violent or psychologically disturbing. Anthropic takes the wellbeing of team members working in this area seriously and provides access to wellness resources and support. Candidates should carefully consider this aspect of the role before applying. Key responsibilities Own the day-to-day operational management of child safety content review workflows, including task routing, queue management, escalation handling, and SLA monitoring Serve as the primary point of contact for review partners conducting child safety content review, including onboarding, training, quality assurance, and ongoing relationship management Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for CSAM, CSEM, and related child safety policy violations Review novel or ambiguous flagged content to drive enforcement decisions and surface policy gaps to the Safeguards policy design team Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in child safety technology, and use these to inform our workflows Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy Minimum qualifications Experience in trust & safety, content moderation operations, or policy enforcement with direct exposure to child safety, CSAM/CSEM, or related child protection harm areas Experience managing or coordinating content review operations, including quality assurance and workflow management Experience standing up and scaling policy enforcement or content review workflows Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends Experience identifying emerging risks and communicating findings to cross-functional stakeholders, such as Product, Policy, Engineering, and Legal teams Understanding of the challenges involved in implementing product policies at scale in the content moderation space Preferred qualifications Deep subject matter expertise in child safety, child sexual exploitation and abuse (CSEA), or online child protection, including familiarity with CSAM/CSEM classification standards (e.g., COPINE, SAM scale) Experience working with or reporting to NCMEC, IWF, or equivalent child safety reporting bodies Familiarity with relevant legal and regulatory frameworks, including CSAM reporting obligations, KOSA, COPPA, or equivalent international frameworks Experience working with generative AI products, including an understanding of how AI systems can be misused to generate or facilitate CSEA Experience designing or evaluating trauma-informed support structures and wellness protocols for content reviewers working with harmful material Proficiency in Python for workflow automation or data analysis Experience working with hash-matching technologies (e.g., PhotoDNA, CSAI Match) or perceptual hashing tools used in CSAM detection Familiarity with age assurance technologies and their role in child safety enforcement Experience in a trust & safety role at a technology company, with an understanding of how platform policy intersects with child protection obligations The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $245,000 — $285,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your initial focus will be account compromise: Anthropic's enforcement systems have to distinguish customers whose accounts have been compromised from actors abusing the platform — and today those two populations can look identical in the data. Stolen credentials and leaked keys are a growing abuse vector, and the collateral damage from enforcement against them lands on legitimate users. You'll own this problem end to end: detection, revocation, user notification, remediation, and the criteria for restoring access — turning what is today ad hoc incident response into a scalable, repeatable program. This position may expand into broader areas of enforcement over time. Safety is core to our mission, and you'll help shape policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way. Key responsibilities Investigate credential-compromise incidents across first-party and third-party platforms, tracing actor behavior across accounts and surfaces Design and operate remediation workflows for compromised accounts: revocation, customer notification, and standards for restoring access Partner with Engineering and Data Science teams to improve how we separate compromised-customer traffic from willful abuse Enforce usage policies with a focus on detecting and mitigating potentially harmful use of AI systems Work with threat intelligence on emerging credential-abuse patterns and the actors behind them Support the Safeguards policy design team by providing detailed feedback on policy gaps based on real enforcement scenarios Keep up to date with emerging AI policy enforcement best practices, and use these to inform our decision-making and workflows Write the policy framework for compromise scenarios, including cases where Anthropic can't independently verify a customer's security posture Act as the enforcement SME when account compromise intersects with active abuse investigations Minimum qualifications Experience in trust and safety, fraud investigation, security operations, or a related field Subject matter expertise in one or more of: account takeover, credential abuse, session security, or incident response Experience designing or operating enforcement, remediation, or customer-recovery flows — not just detection Comfort using data (SQL or similar tools) to trace actor behavior across accounts and to measure what's working A thoughtful perspective on the tension between protecting the platform and restoring access for legitimate compromised customers Strong written communication skills, with experience producing clear briefs about messy incidents for technical and non-technical stakeholders Excellent judgment and the ability to collaborate with team members while navigating rapidly evolving priorities and workstreams Preferred qualifications Familiarity with credential-theft ecosystems or threat intelligence tooling Experience working with fraud/risk vendors (e.g., device or network intelligence platforms) Experience with API platform abuse specifically, not just consumer-account ATO A deep interest in AI safety and responsible technology development Experience writing effective prompts for generative AI systems in a content review or enforcement context The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $245,000 — $285,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Analyst on the account abuse team, you'll build and execute enforcement workflows that keep our products safe, with a focus on detecting and mitigating potential harm. Your focus will be driving a number of enforcement areas including access controls and identity verification. You'll own the policy layer of these systems: what we ask users for, when, on what grounds, and what passes. The work sits at the intersection of policy, operations, and regulatory exposure. This position may expand into broader areas of enforcement over time. Safety is core to our mission, and you'll help shape policy enforcement so that our users can safely interact with and build on top of our products in a harmless, helpful, and honest way. Key responsibilities Set collection policy for identity signals: what we request, in which enforcement states, and what evidence reinstates access Improve verification program health — accuracy, appeal coverage, and consistency of outcomes Own access policy for hard cases, including geographic restrictions and reseller arrangements Run the operational queue for access-control cases alongside contractor support, authoring playbooks and QA'ing scaled review output Work with Legal, Public Policy, and Privacy stakeholders to keep our approach proportionate, privacy-preserving, and responsive to an evolving regulatory landscape Coordinate enforcement consistency with third-party platform partners Keep up to date with emerging AI policy enforcement best practices, and use these to inform our decision-making and workflows Stand up graduated enforcement in practice — verification requests, conditional reinstatement, and appeal pathways that satisfy regulatory requirements for automated decisions Minimum qualifications Experience in trust & safety, integrity, or risk policy work Hands-on operational experience — you've owned or quality-checked live enforcement queues, not only authored policy Subject matter expertise in one or more of: KYC, identity verification, age or identity assurance, or verification program operations Experience navigating evolving regulatory landscapes (including frameworks like the DSA and GDPR) as design constraints rather than blockers Experience driving cross-functional initiatives with Product, Engineering, Legal, and Policy partners — especially where safety, privacy, and usability tradeoffs need to be navigated together Comfort using data (SQL or similar tools) to measure what's working and inform decisions Strong written communication skills, with experience producing clear briefs and recommendations for technical and non-technical stakeholders Excellent judgment and the ability to make consistent, defensible calls on ambiguous cases Preferred qualifications Experience with graduated/tiered enforcement systems rather than binary ban models Experience with geographic access restrictions, sanctions screening, or reseller/channel policy Experience building or scaling a contractor review bench A deep interest in AI safety and responsible technology development Experience writing effective prompts for generative AI systems in a content review or enforcement context The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $285,000 — $330,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As an Enforcement Analyst focused on Bio Harms, you will play a critical role in protecting against the misuse of AI systems for biological and related CBRNE harms. You will enforce our Usage Policy with a specific focus on detecting and mitigating bio risks, investigating potential violations, and help continuously strengthen our safeguards. The work sits at the intersection of biosecurity threat analysis and platform enforcement: you will read real model interactions and make fast, well-reasoned calls about whether activity is benign research or a credible attempt at harm. This role is a fit for someone who understands the dual-use nature of biology and enabling technologies well enough to separate the benign from the malicious. You will own and continuously improve the enforcement monitoring workflows for the bio-harms area, and you will work closely with Policy, Threat Intelligence, Data Science, and Engineering cross-functional partners to accomplish tasks at scale. Safety is core to our mission, and your work will directly protect individuals, communities, and critical systems. Important context for the role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including material of a sexual, violent, or psychologically disturbing nature. The role also carries a shared on-call responsibility across the Policy and Enforcement teams. Key responsibilities Enforce Usage Policies with a specific focus on detecting and mitigating potential bio risks and harmful use of AI systems. Take ownership of enforcement monitoring workflows for the bio-harms area, improving end-to-end detection, investigation, triage, and escalation processes. Monitor and analyze platform activity to identify emerging patterns related to biological (and adjacent chemical, radiological, nuclear, and explosive) threats that may require policy updates or enforcement action. Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy across a technically complex content surface. Conduct thorough investigations of potential violations, gathering and documenting evidence to support enforcement decisions. Proactively surface trends and propose improvements to detection methods and review workflows. Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for bio-related policy violations. Partner with Policy and Threat Intelligence groups to understand potential exploits and contribute to risk-assessment frameworks, and partner with engineers iterating on safety systems. Provide enforcement-grounded feedback on policy gaps, and handle escalations and time-sensitive situations related to potential bio-related Usage Policy violations. Minimum qualifications Hold a degree in a bio-related field (e.g., microbiology, molecular biology, biochemistry, public health) and/or relevant professional experience in a related field. Possess experience in Trust & Safety, content moderation, or policy enforcement at platform scale, including working with generative AI tools to refine and optimize content review and enforcement workflows. Possess experience in utilizing AI tools to develop data dashboards for metrics collection in support of continuous improvement efforts. Can analyze complex, ambiguous situations and make well-reasoned, defensible decisions under time pressure. Are proactive and self-directed: you spot trends, dig in, and ship improvements on your own initiative. Communicate clearly in writing and can translate technical bio concepts for diverse audiences, including both technical and non-technical stakeholders. Preferred qualifications Subject matter expertise in biodefense, biosecurity, WMD/CBRNE non-proliferation, or threat-intelligence. An understanding of where real-world biological risk actually lies — adversary intent, acquisition and weaponization pathways, and dual-use concerns, including familiarity with cross-platform threat analysis and open-source intelligence (OSINT) techniques relevant to weapons of mass destruction. Proficiency in SQL and/or other data analysis tools to draw insights from large datasets and monitor enforcement workflow health. Familiarity with dual-use research of concern (DURC), select agents, and relevant policy frameworks. The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $245,000 — $285,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Lead on the User Well-Being team, you will be responsible for managing our child safety, mental health, abuse and exploitation, and age assurance enforcement workflows. This will include managing the team responsible for scaling, maintaining, and continuously improving the systems and processes we use to detect and respond to these harms. This is a management role. You will serve as the central point of contact for those who conduct content review, managing day-to-day workflows, quality assurance, and escalation processes. You will also work closely with internal Engineering, Policy, and Legal teams to scale detection systems and close enforcement gaps as the threat landscape evolves. Important context for this role: In this position you will regularly be exposed to and engage with explicit content of a sexual nature involving minors, as well as content that may be violent or psychologically disturbing. Anthropic takes the wellbeing of team members working in this area seriously and provides access to wellness resources and support. Candidates should carefully consider this aspect of the role before applying. Key responsibilities Manage a team of individual contributors across multiple policy areas under the User Well-Being banner Serve as the primary point of contact for review partners conducting content review, including onboarding, training, quality assurance, and ongoing relationship management Design and improve enforcement workflows to scale effectively as volume grows, while maintaining high accuracy and consistency across review decisions Partner with Engineering and Data Science teams to optimize detection models and automated enforcement systems for User Well-Being policies Develop and maintain internal documentation, decision trees, and review guidelines that enable accurate and consistent enforcement at scale Keep up to date with emerging AI policy enforcement best practices, evolving legal frameworks, and developments in technology, and use these to inform our workflows Identify and report trends in misuse patterns to internal stakeholders, including Policy, Legal, and Trust & Safety leadership Coordinate reporting obligations to relevant external bodies (e.g., NCMEC) in accordance with applicable law and Anthropic policy Minimum qualifications Experience managing teams in the User Well-Being space Experience in trust & safety, content moderation operations, or policy enforcement with direct exposure to child safety, mental health, abuse and exploitation, and age assurance harm areas Experience managing or coordinating content review operations, including quality assurance and workflow management Experience standing up and scaling policy enforcement or content review workflows Proficiency in SQL and/or other data analysis tools to monitor workflow health, review queue metrics, and surface enforcement trends Experience identifying emerging risks and communicating findings to cross-functional stakeholders, such as Product, Policy, Engineering, and Legal teams Understanding of the challenges involved in implementing product policies at scale in the content moderation space Preferred qualifications Deep subject matter expertise in child safety, child sexual exploitation and abuse (CSEA), online child protection, mental wellness, and age assurance Experience working with or reporting to NCMEC, IWF, or equivalent child safety reporting bodies Familiarity with relevant legal and regulatory frameworks, including CSAM reporting obligations, KOSA, COPPA, or equivalent international frameworks Experience working with generative AI products, including an understanding of how AI systems can be misused to generate or facilitate abusive content Experience designing or evaluating trauma-informed support structures and wellness protocols for content reviewers working with harmful material Proficiency in Python for workflow automation or data analysis Experience working with hash-matching technologies (e.g., PhotoDNA, CSAI Match) or perceptual hashing tools used in CSAM detection Familiarity with age assurance technologies and their role in child safety enforcement The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $285,000 — $330,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.