Senior Applied Scientist, Efficient LLM Inference & Model Optimization
In short
- →Científico aplicado senior que optimiza inferencia de LLM/VLM con impacto real en producción.
- →Día a día: investiga, pruebas, codifica prototipos y los entrega a ingenieros para implementar.
- →Lo destacable: el trabajo se publica, se comparte y se despliega, no solo se queda en un paper.
Proficiency in English required for collaboration and documentation.
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
The questions they'll ask you
1. ¿Cómo diseñarías un experimento para evaluar el impacto de la compresión del KV-cache en latencia y calidad?
No card. Upload your resume and the full dossier is ready in ~1 minute.
💵 USD · Remote · No visa
Not finding what you want? Try Micro1
Micro1 places engineers directly at US companies paying in USD. One vetting, multiple offers — no cold applying.
What they ask for
- ✓PhD en ciencia de datos, inteligencia artificial o área relacionada.
- ✓Experiencia comprobada en optimización de inferencia de modelos grandes.
- ✓Habilidades sólidas en PyTorch, CUDA y herramientas de bajo nivel.
- ✓Capacidad para diseñar experimentos rigurosos con métricas medibles.
- ✓Experiencia colaborando con MLEs y equipos de ingeniería en producción.
- ✓Publicaciones en conferences relevantes (NeurIPS, ICML, ACL, etc.)
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Nebius?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
More jobs like this
Looking for something similar?
Leave your email and we'll alert you when matching jobs appear.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.