Staff Software Engineer, Inference
In short
- →Ingeniero de software de alto impacto en plataformas de inferencia de IA a gran escala.
- →Diseñas arquitecturas, optimizas latencias y lideras iniciativas técnicas en Kubernetes y GPU.
- →Destaca por liderar mejoras en rendimiento con enfoque en P99 y costos por token.
Proficiency in English required for collaboration across global teams.
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
The questions they'll ask you
1. ¿Cómo has optimizado la latencia P99 en un sistema de inferencia con múltiples modelos concurrentes?
No card. Upload your resume and the full dossier is ready in ~1 minute.
💵 USD · Remote · No visa
Not finding what you want? Try Micro1
Micro1 places engineers directly at US companies paying in USD. One vetting, multiple offers — no cold applying.
What they ask for
- ✓8+ años en sistemas distribuidos o plataformas en la nube.
- ✓Experiencia comprobada liderando iniciativas técnicas a gran escala.
- ✓Dominio de Go, Python o C++.
- ✓Experto en Kubernetes a escala productiva.
- ✓Experiencia práctica en sistemas de inferencia (batching, caché, memoria, precisión mixta).
- ✓Capacidad demostrada de mejorar latencias de cola (P95/P99) con métricas.
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Coreweaveu?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
Looking for something similar?
Leave your email and we'll alert you when matching jobs appear.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.