ApplySarthi

Research Engineer - Eval Platform

Mistral.ai

Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.

Got this interview? Our apps help you get the job.

Skills named in this job

Read from the description itself, not inferred.

This role on the market

1,227 open research roles across 215 companies are on ApplySarthi right now, most of them in Bengaluru (32), Mumbai (18), Hyderabad (17).

What research roles keep asking for: Python (29%), Machine learning (27%), PyTorch (14%), LLMs (14%) — counted across their open postings here.

Remote Research Engineer jobs · CI/CD jobs · Kubernetes jobs · LLMs jobs · Python jobs

Mistral.ai has 25 open roles listed here.

Counted across 14 company job boards, updated as roles open and close.

Preparing for this interview

Interviews for research roles keep coming back to Python, Machine learning, PyTorch, LLMs. Practise those questions before you sit with Mistral.ai.

Questions you are likely to be asked

  1. Why do you want to join Mistral.ai?
  2. What is your experience with CI/CD? Tell me one thing you learned the hard way.
  3. Tell me about a hard bug you tracked down. How did you find the cause?
  4. How do you decide what to test, and what does good code review look like to you?
  5. Describe a time a deadline forced a trade-off in quality. What did you choose and why?

Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.

Practise the Research Engineer - Eval Platform at Mistral.ai interview free →

About Mistral Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms. We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited. The Role Evaluation is how we decide which models, checkpoints and recipes ship. As a Research Engineer on the Eval Platform team, you will build the infrastructure every science team relies on to measure model quality, and make it reliable, reproducible and fast. You don't need to have designed benchmarks before. You do need to care about what a score means, and about when a difference between two runs is real. What you will do Build systems that keep eval results reproducible and comparable over time, as models, benchmarks and code evolve. Run evaluations at scale across our GPU clusters, from model serving to scoring. Make eval results easy to access, explore and trust, through APIs and dashboards that researchers use every day. Catch broken or noisy evals before they mislead research decisions. Support evaluation of agentic, multi-turn and tool-using models. Work closely with researchers to turn new evaluation needs into robust, shared tooling. What we're looking for Master's or PhD in Computer Science, or equivalent experience. 4+ years building production-grade software, ideally large-scale ML codebases or distributed systems. Excellent Python and strong software-design instincts: testing, code review, CI/CD. Experience running workloads on GPU clusters (Slurm, Kubernetes, Ray or similar). Familiarity with LLM inference and evaluation. A product mindset: researchers are your users. Self-starter, low-ego, collaborative. Nice to have Experience building or maintaining evaluation harnesses or benchmarks. Hands-on experience with inference engines such as vLLM or SGLang. Experience with agentic or RL environments. Statistics for experimentation: variance estimation, significance testing. Open-source contributions to ML tooling. What We Offer We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks. For the most up-to-date details on benefits available in your location, please refer to our Benefits page . Privacy Policy Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy . Find more English Speaking Jobs in France on Arbeitnow

Match this job to your CV

ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.

Check my match →

Similar open roles

Need answers during your interview? Try Live Sarthi.

Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.

Try Live Sarthi free →

A Windows app, from the same team as ApplySarthi.

Listed on arbeitnow · posted 2026-10-10. ApplySarthi collects openings and links to application pages; the role is advertised by Mistral.ai, not by us.