Performance Engineer, Inference Engine
Anthropic · San Francisco, CA | New York City, NYmid
In short
- ▸Ingeniero de rendimiento enfocado en optimizar el motor de inferencia de LLM a escala de millones de usuarios.
- ▸Trabajas en hardware, memoria, comunicación entre dispositivos y coordinación distribuida para maximizar eficiencia y mantener calidad del modelo.
- ▸Destaca que el sistema es interno, crítico para la seguridad y escalabilidad de Claude.
Fluency in English is required for the role.
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →What they ask for
- ✓Modelo mental sólido de la inferencia de LLM (prefill, decode, aceleradores, memoria, interconexión)
- ✓Capacidad probada de aprender rápido y entregar cambios significativos en sistemas complejos
- ✓Programación de sistemas sólida en Rust, C++ o similar con enfoque en calidad y pruebas
- ✓Enfoque analítico: observar, modelar, probar y cambiar iterativamente
- ✓Bajo ego, disposición a preguntar, aceptar feedback y colaborar
- ✓Gusto por programar en pareja y compromiso con el impacto social del trabajo
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Anthropic?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.