InterviewHack.ai
Start free
Jobs / Mistral.ai

Research Engineer - Eval Platform

Mistral.ai·Paris
Apply on company site ↗Share on WhatsApp
✓ Free to start✓ Runs in your browser✓ First dossier, no card✓ Ready in ~1 minute

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Your first dossier is free.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

💵 USD · Remote · No visa

Not finding what you want? Try Micro1

Micro1 places engineers directly at US companies paying in USD. One vetting, multiple offers — no cold applying.

Get matched by Micro1 →
📬Jobs picked for YOUR resume, every morning on WhatsApp. Free: text “vacantes” and the bot sends your daily matches. Subscribe →

Who should you write to at Mistral.ai?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

About Mistral Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms. We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited. The Role Evaluation is how we decide which models, checkpoints and recipes ship. As a Research Engineer on the Eval Platform team, you will build the infrastructure every science team relies on to measure model quality, and make it reliable, reproducible and fast. You don't need to have designed benchmarks before. You do need to care about what a score means, and about when a difference between two runs is real. What you will do Build systems that keep eval results reproducible and comparable over time, as models, benchmarks and code evolve. Run evaluations at scale across our GPU clusters, from model serving to scoring. Make eval results easy to access, explore and trust, through APIs and dashboards that researchers use every day. Catch broken or noisy evals before they mislead research decisions. Support evaluation of agentic, multi-turn and tool-using models. Work closely with researchers to turn new evaluation needs into robust, shared tooling. What we're looking for Master's or PhD in Computer Science, or equivalent experience. 4+ years building production-grade software, ideally large-scale ML codebases or distributed systems. Excellent Python and strong software-design instincts: testing, code review, CI/CD. Experience running workloads on GPU clusters (Slurm, Kubernetes, Ray or similar). Familiarity with LLM inference and evaluation. A product mindset: researchers are your users. Self-starter, low-ego, collaborative. Nice to have Experience building or maintaining evaluation harnesses or benchmarks. Hands-on experience with inference engines such as vLLM or SGLang. Experience with agentic or RL environments. Statistics for experimentation: variance estimation, significance testing. Open-source contributions to ML tooling. What We Offer We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks. For the most up-to-date details on benefits available in your location, please refer to our Benefits page . Privacy Policy Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy . Find more English Speaking Jobs in France on Arbeitnow

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsCompanies hiringAll free toolsResume verdict (Jev)Free cover letterInterview questions by roleTechnical assessment simulator"Tell me about yourself" answerFree ATS checkerInterview-English checkSalary checkSalary negotiation scriptFree STAR answerLinkedIn headline + AboutLATAM salary reportFree coursesBlogTailored CVSpoken practicePricingAffiliates — 30%

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

HRBP

Mistral.ai · Paris

→

Talent Acquisition Specialist, Early Careers - EMEA

Mistral.ai · Paris

→

Applied AI, Technical Lead, Forward Deployed AI Engineer - London

Mistral.ai · London

→

Applied AI, Forward Deployed Machine Learning Engineer - London

Mistral.ai · London

→