InterviewHack.ai
Start free
Jobs / Anthropic

Engineering Manager, Safeguards

Anthropic · London, UKlead

In short

  • ▸Liderazgo de un equipo que construye herramientas para revisar y asegurar el uso seguro de modelos de IA.
  • ▸Desarrollas plataformas para investigar riesgos, con enfoque en privacidad y automatización con Claude.
  • ▸Destacado: usas IA (Claude) para ampliar el trabajo humano en revisiones de contenido, manteniendo control humano.

Excellent communication skills, including the ability to explain technical tradeoffs to no

Apply on company site ↗Share on WhatsApp

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.

What they ask for

  • ✓Experiencia gestionando equipos de ingeniería de software.
  • ✓Formación técnica en ingeniería de plataformas o full-stack.
  • ✓Habilidades comprobadas para crear herramientas internas con usuarios operativos exigentes.
  • ✓Experiencia en trabajo colaborativo con equipos no técnicos (política, legal, operaciones).
  • ✓Capacidad para comunicar trade-offs técnicos a stakeholders no técnicos.
  • ✓Compromiso con el impacto social de la IA y la seguridad de los sistemas.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

Review ToolingAnalytics capabilitiesPrivacy-preserving primitivesData retention commitmentsSandbox environmentClaudeEnforcement workflowsInternal toolsPlatform engineeringCross-functional collaboration

Who should you write to at Anthropic?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role The Safeguards team is responsible for ensuring our models and products are developed and deployed safely. We're looking for an Engineering Manager to lead our Review Tooling team, which builds the systems that humans — and increasingly Claude — use to investigate potential harms and take enforcement actions across Anthropic's first-party products and third-party cloud platforms. This is a foundational role: you'll own the tools our safety investigators rely on to understand what's happening on our platforms and act on it, as well as the platform underneath those tools. That platform includes analytics capabilities, privacy-preserving primitives that keep review workflows compatible with our data retention commitments, and a sandbox environment where new review interfaces and workflows can be built and iterated quickly. As model capabilities and usage grow, you'll also drive how we scale review through automation — building systems where Claude meaningfully extends what human reviewers can do, while keeping people in the loop where their judgment matters most. You'll partner closely with policy, operations, data science, and legal teams to ensure our enforcement systems are effective, accurate, and trustworthy. Key responsibilities • Lead, grow, and develop a team of engineers building investigation, review, and enforcement tooling for both first-party and third-party platform surfaces • Define the vision and roadmap for our review tooling platform, including analytics, privacy-compatible data access primitives, and a sandbox for rapidly developing new review interfaces • Drive the team's strategy for scaling review through automation, including enabling reviewers to use Claude effectively and building toward Claude-assisted and Claude-driven review workflows • Partner with policy, operations, legal, privacy, and data science stakeholders to translate enforcement and investigation needs into reliable, well-designed systems • Ensure review tooling evolves alongside new privacy primitives and data retention commitments, so reviewers can do their work without compromising user trust • Create clarity for the team and stakeholders in an ambiguous and evolving environment • Take an inclusive, equitable approach to hiring and coaching top technical talent, and maintain a high-performing team • Contribute to engineering-wide initiatives as a member of Anthropic's engineering management community Minimum qualifications • Experience managing software engineering teams, including hiring, coaching, and developing engineers • A technical background in full-stack or platform engineering, with the ability to engage deeply in architecture and design discussions • Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably • Experience working cross-functionally with non-engineering partners such as operations, policy, or legal teams • Excellent communication skills, including the ability to explain technical tradeoffs to non-technical stakeholders • Care about the societal impacts of AI and want your work to make powerful systems safer Preferred qualifications • 4+ years of management experience, 10+ years of industry software engineering experience • Experience building trust and safety, integrity, fraud, or abuse-prevention tooling, or other systems supporting hum

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsFree ATS checkerInterview-English checkSalary checkLATAM salary reportFree coursesBlogTailored CVSpoken practiceIt's free

Remote jobs

ReactPythonFull-StackLATAMMexicoSee all →

Prepare

Practice with a coachFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Mega Account Executive, Meta

Anthropic · San Francisco, CA | New York City, NY

→

Communications Lead, France and Southern Europe

Anthropic · Paris, France

→

Manager, APAC Recruiting

Anthropic · Sydney, Australia

→

Staff Software Engineer, Observability & Profiling

Anthropic · London, UK

→