InterviewHack.ai
Empezar gratis
Vacantes / Anthropic

Safeguards Enforcement Analyst, Conventional Weapons

Anthropic · New York City, NY; Remote-Friendly (Travel-Required) | San Francisco, CA | Washington, DCRemotomid

En corto

  • ▸Analista encargado de detectar y evitar el mal uso de IA para crear o facilitar armas convencionales.
  • ▸Diseña sistemas automatizados, evalúas contenido peligroso y colaboras con ingenieros para mejorar la seguridad.
  • ▸Trabajas con contenido gráfico y violento, lo que requiere resiliencia emocional y experiencia en amenazas reales.

Proficiency in English

Postularme en la empresa ↗Compartir por WhatsApp

En ~1 minuto te damos: quién te entrevista, las preguntas probables con respuestas desde tu CV, y tu CV adaptado a esta vacante. Gratis, sin tarjeta.

¿Qué piden?

  • ✓Experto aplicado en sistemas de armas convencionales y capacidad para evaluar riesgos técnicos.
  • ✓Experiencia en aplicación de políticas, inteligencia de amenazas o contraterrosmismo con exposición a contenido dañino.
  • ✓Capacidad para escalar flujos de revisión de contenido y crear evaluaciones efectivas.
  • ✓Dominio de SQL o herramientas de análisis de datos para monitorear sistemas de cumplimiento.
  • ✓Experiencia identificando riesgos emergentes y comunicando hallazgos a equipos multidisciplinarios.
  • ✓Conocimiento previo de productos de IA generativa y uso de prompts para revisión de contenido.

¿No cumplís todo? Es lo normal — tu dossier gratis te dice qué gaps tenés y cómo cubrirlos en la entrevista.

SQLData analysis toolsGenerative AI productsPolicy enforcement workflowsAutomated detection systemsEnforcement evalsReviewer documentationThreat intelligence platformsRegulatory frameworksProduct policy implementation

¿A quién escribirle en Anthropic?

Tu dossier gratis identifica a las personas que te entrevistarían — con su background, qué valoran y cómo escribirles para destacar antes de aplicar.

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role As a Safeguards Enforcement Analyst focused on Conventional Weapons, your work spans detecting and mitigating attempts to misuse Anthropic's AI systems to facilitate real-world harm, specifically utilizing conventional weapons and dangerous technology. You will be responsible for building and executing operational workflows to assess model behavior, drive enforcement decisions, and develop evals across a technically demanding range of policy areas. Important context for this role: In this position you may be exposed to and engage with explicit content spanning a range of topics, including those of a violent, graphic, hateful, or psychologically disturbing nature. Key responsibilities • Design and architect automated enforcement systems and review workflows that scale effectively while maintaining high accuracy • Develop and maintain evals that measure model performance on these policy areas, surface regressions, and inform policy and model improvements • Partner with Engineering and Data Science to optimize detection and automated enforcement systems for potential policy violations • Review flagged content to drive enforcement decisions and surface policy gaps, with particular attention to novel or technically sophisticated misuse attempts + emerging tactics • Support the Safeguards policy design team by providing structured feedback on policy gaps and enforcement ambiguities based on real enforcement scenarios • Develop and maintain enforcement guidelines and reviewer documentation that enable accurate, consistent enforcement across a wide range of content • Keep up to date with emerging weapons trends and applications, regulatory changes, and AI policy enforcement best practices, and apply these to inform our workflows and evals • Identify and escalate emerging misuse patterns, novel attack vectors, and signs of coordinated violent activity Minimum qualifications • Have deep, applied expertise in weapons systems and can translate complex technical evidence to make enforcement decisions • Experience in policy enforcement, threat intelligence, counterterrorism, government, or a closely related field, with direct exposure to harmful content, dangerous technology, or physical harm facilitation • Experience standing up and scaling policy enforcement or content review workflows • Proficiency in SQL and/or other data analysis tools to draw insights from large datasets and monitor enforcement workflow health • Experience identifying emerging risks and threat actors, and communicating findings to a diverse set of stakeholders, such as Product, Policy, Engineering, and Legal teams • Experience working with generative AI products, including writing effective prompts for content review and enforcement • Understanding of the challenges involved in implementing product policies at scale, including in the content moderation space Preferred qualifications • Subject matter expertise in conventional weapons and dangerous technology, autonomous systems, or critical infrastructure protection • Familiarity with relevant legal and regulatory frameworks governing dangerous technology, conventional weapons, and critical infra

No apliques sin prepararte

Investigamos quién te entrevista, adaptamos tu CV y te ensayamos en vivo — gratis la primera.

InterviewHack.ai

Preparate para la entrevista exacta: quién te entrevista, tu CV a medida y coach real.

Producto

VacantesRevisar CV (ATS) gratis¿Cómo suena tu inglés?¿Te pagan bien?Reporte de sueldos LATAMCursos gratisBlogCV a medidaPráctica habladaEs gratis

Empleos remotos

ReactPythonFull-StackLATAMMéxicoVer todas →

Preparate

Practicá con coachFrontendBackendAI EngineerPor empresaVendete con tu CV

Empresa

Buscás talentoAcerca deContactoPrivacidadTérminos

© 2026 InterviewHack.ai · Tu CV es tuyo. Nunca se usa para entrenar nada. · Un producto de IA-PTY

Vacantes similares activas

Mega Account Executive, Meta

Anthropic · San Francisco, CA | New York City, NY

→

Engineering Manager, Safeguards

Anthropic · London, UK

→

Communications Lead, France and Southern Europe

Anthropic · Paris, France

→

Manager, APAC Recruiting

Anthropic · Sydney, Australia

→