Senior ML Engineer (Token Factory)
Nebius · Berlinsenior
In short
- ▸Ingeniero de ML sénior enfocado en optimizar inferencia y fine-tuning de modelos grandes a escala en GPU.
- ▸Días a día: perfilar trabajo en GPU, diseñar pipelines de baja precisión y mejorar motores de inferencia como vLLM o TensorRT-LLM.
- ▸Destacado: trabajar en una de las mayores nubes de GPU del mundo, con miles de GPUs y un impacto directo en el rendimiento de modelos como GPT-OSS o GLM-5.
Excellent command of the English language, alongside superior writing, articulation, and c
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →What they ask for
- ✓Conocimiento profundo de arquitecturas de transformadores y ML teórico.
- ✓Experiencia con herramientas de perfilado de GPU (Nsight, PyTorch profiler).
- ✓Entendimiento de jerarquía de memoria y tradeoffs de compute/memory en GPU.
- ✓Familiaridad con conceptos clave de LLM (MHA, RoPE, KV-cache, Flash Attention, cuantización).
- ✓Experiencia con frameworks modernos de deep learning y alta habilidad en ingeniería de software (Python, CI/CD, testing).
- ✓Habilidades comunicativas y de liderazgo sólidas.
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Nebius?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.