ML Systems Engineer — Inference Acceleration
Arago · Paris Officessenior
In short
- ▸Ingeniero de sistemas de ML que optimiza inferencia en un acelerador óptico-CMOS único.
- ▸Trabaja en kernels, ejecución distribuida, servidor de inferencia y optimización de memoria en un stack de software en desarrollo.
- ▸Destaca por ser el único equipo que trabaja con un procesador híbrido óptico-CMOS con rendimiento de orden de magnitud superior.
Proficient level of English
In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.
🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →What they ask for
- ✓Experiencia sólida en inferencia de alto rendimiento de ML.
- ✓Conocimiento profundo de arquitectura de computadoras y modelos de ejecución en aceleradores.
- ✓Experiencia con CUDA, Triton, ROCm/HIP o entornos equivalentes para programación de kernels.
- ✓Habilidades sólidas en C++ y Python para desarrollo en stack personalizado.
- ✓Experiencia con sistemas de inferencia modernos como vLLM, SGLang o TensorRT-LLM.
- ✓Nivel de inglés proficient, con trabajo en entornos técnicos en inglés.
Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.
Who should you write to at Arago?
Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.
Don't apply unprepared
We research who's interviewing you, tailor your CV and rehearse you live — first one free.