InterviewHack.ai
Start free
Jobs / Anthropic

Product Manager, Safeguards (Account Integrity & Abuse)

Anthropic·San Francisco, CA | New York City, NYmid

In short

  • →Gestionar sistemas de seguridad para proteger a usuarios de abusos en IA avanzada.
  • →Trabajar con investigación, ingeniería y políticas para crear defensas técnicas y evaluaciones de riesgo.
  • →Destacar por liderar productos desde cero en un entorno altamente incierto y rápido.

Ability to write safety evals and communicate externally about safety.

Apply on company site ↗Share on WhatsApp

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

What they ask for

  • ✓Experiencia en diseño y construcción de sistemas de seguridad para IA.
  • ✓Capacidad para tomar decisiones técnicas con equipo multidisciplinario.
  • ✓Habilidades sólidas en métricas y evaluación de riesgos.
  • ✓Capacidad de priorizar en entornos ambiguos y cambiantes.
  • ✓Comprensión profunda del uso de productos y sus riesgos de seguridad.
  • ✓Experiencia en lanzar productos desde cero en entornos dinámicos.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

AI safety systemssafety evaluationsrisk metricsproduct strategycross-functional collaborationAI/ML researchsoftware engineeringpolicy alignmentuser experience (UX)deployment risk mitigation

Who should you write to at Anthropic?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role Anthropic is dedicated to developing AI assistants that are helpful, harmless, and honest. As usage of our AI services grows, we need to ensure they are not misused. The Safeguards team is at the forefront of protecting our users from the risks of powerful AIs as well as ethical, technical, and social risks from the use of generative AI. The Safeguards team at Anthropic builds protections for new AI features enabled by the research teams and protects new products and surfaces developed by our product teams. As a Product Manager for the Safeguards team at Anthropic, you will own the ideation, design, development and deployment of Safeguards systems and relevant product UX to ensure we are advancing frontier models safely to users across various cloud platforms. You will work closely with our research and product teams to develop detections, evals, interventions, and tools to measure and mitigate deployment and user risks. We are looking for a product manager who is deeply committed to making AI safe and beneficial for humanity. You are aware of the risks and are committed to working with experts and coming up with ideas for Anthropic to implement. You have deep technical expertise in development, deployment and measurement of Safeguards systems. You thrive in rapidly moving and ambiguous environments. Responsibilities • Determine how to build in safety by design upstream and leverage downstream defenses for Anthropic’s frontier models, AI products, customers on different surfaces - Claude.ai, 1P API, external Cloud providers. • Ability to write safety evals and communicate externally about safety. • Drive impact via ruthless prioritization by clearly defining problems, solution options forward, clarity on both business & technical tradeoffs and accordingly clear requirements toward MVP vs. ideal state. • Align & collaborate with policy, enforcement, research, engineering and cross functional stakeholders. • Understand the AI landscape and ecosystem to plan for mitigation of deployment risks of increasingly powerful models and determined adversaries. • Lead the development of metrics to understand the area, performance, blindspots to help inform future project planning. Minimum Qualifications You may be a good fit if you have: • Ability to make technical tradeoff decisions; ideally with experience working across policy experts, AI/ML research engineers and software engineering teams to design and build state of the art safety systems. • Strong user understanding of how our products are used, their Safeguards concerns and how we provide the best solutions. • Demonstrated ability to build product and engineering strategy across multiple cross-functional teams for a rapidly changing space. • Demonstrated experience in designing and building metrics to evaluate risks, system performance, user impact and making crisp tradeoffs • Very strong ability to navigate, and prioritize amidst rapidly changing product specs, and to flex into different domains to bring clarity and execute. • Evidence of exercising judgment and decision making in ambiguous situations. • Planning, building, launching and measuring new products / systems in a zero to one environment. • Ability to

More jobs like this

Remote Product Manager jobs

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsCompanies hiringFree ATS checkerInterview-English checkSalary checkLATAM salary reportFree coursesBlogTailored CVSpoken practiceIt's free

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Senior Manager, IT SOX

Anthropic · San Francisco, CA

→

Software Engineer, Beneficial Deployments

Anthropic · San Francisco, CA | New York City, NY

→

Applied AI Architect

Anthropic · Mumbai, India

→

Customer Success

Anthropic · Bangalore, India

→