Research Operations, Reinforcement Learning
In short
- →Eres el operador de alto impacto para el equipo de Aprendizaje por Refuerzo de Anthropic, asegurando que los líderes se enfoquen en lo más importante.
- →Tu día a día: gestionas tiempo de líderes, priorizas decisiones, simplificas comunicación, y anticipas problemas antes de que bloquen el avance.
- →Lo destacado: tienes acceso directo a la alta dirección del equipo más técnico de la compañía, influyendo en el rumbo de modelos como Claude.
Fluent written and spoken English required
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →What they ask for
- ✓Compromiso profundo con la misión de Anthropic
- ✓Fluidez técnica para entender discusiones sobre entrenamiento de modelos RL
- ✓Capacidad de escribir claro y conciso, transformando material técnico en acciones ejecutables
- ✓Habilidad para interpretar dashboards y mejorar visualizaciones de datos
- ✓Atención extrema al detalle y seguimiento exhaustivo de compromisos
- ✓Capacidad de anticipar problemas y bloqueos antes de que ocurran
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Anthropic?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
Looking for something similar?
Leave your email and we'll alert you when matching jobs appear.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.