You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
Mistral AI

Research Engineer, Forge

  • Hybrid

Salary

Not stated

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms.

We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited.

Role summary

As a Research Engineer on Forge, you will turn real customer requirements into reliable training and deployment workflows. You’ll work end‑to‑end across model adaptation and post‑training (CPT/SFT/RL/distillation), evaluation, data, and infrastructure. The role bridges research experimentation and production constraints.

This role sits in Applied Science, with direct impact on client outcomes. You’ll collaborate closely with scientists, engineers, product, and customer‑facing teams to ensure Forge projects ship, are maintainable, and can be trusted by others.

Interview focus can vary (algorithms, infrastructure, evals, or data). You don’t need to match every bullet below to apply.

Responsibilities

  • Build and improve post‑training and evaluation workflows (CPT/SFT/RL/distillation), turning prototypes into repeatable Forge “recipes”.
  • Develop tools and pipelines for synthetic data generation, data curation, training, evaluation, and deployment.
  • Debug and harden large‑scale ML systems: distributed training, scheduling/execution, checkpointing, observability, and reproducibility.
  • Improve the Forge codebase via clear APIs, tests, documentation, and maintainable abstractions.
  • Push the frontier of our RL training stack (e.g., high-throughput async rollout and scalable post‑training systems at frontier-model scale)
  • Make sure Forge deployment is seamless and adaptable to a diversity of clients (hardware access, software stack, cloud and on-premises, …)
  • Partner with researchers and infrastructure engineers to translate bottlenecks into concrete system improvements.

Requirements

  • Strong Python engineering skills and experience working in large codebases (testing, code review, CI, operational ownership).
  • Hands‑on experience with PyTorch, JAX, or similar.
  • Strong systems and infrastructure fundamentals.
  • Experience with LLM training or post‑training: fine‑tuning, RL, distillation, evaluation, and/or data pipelines.
  • Excellent debugging skills in ambiguous systems (distributed jobs, data issues, quality regressions, infra failures).
  • Clear communication with technical and non‑technical stakeholders.
  • High agency, low ego, and comfort in fast‑moving, under‑specified environments.
  • Distributed training experience (FSDP, DeepSpeed, Megatron, etc.).
  • Cluster/orchestration experience (SLURM, Ray, Kubernetes, Kueue, Karpenter, Skypilot, etc.).
  • Experience building reliable ML infrastructure, evaluation systems, or large‑scale data processing pipelines.
  • Research experience in LLMs, agents, multimodal models, reasoning, code, or domain adaptation.
  • Open‑source contributions, publications, or widely used internal tooling.
  • Experience training multi‑billion‑parameter models (pre‑training or RL). Experience training on petabyte- and exabyte-scale datasets.
  • Ability to identify bottlenecks across the stack and drive improvements from first principles.

Benefits

  • We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.
  • For the most up-to-date details on benefits available in your location, please refer to our Benefits page .
  • Privacy Policy
  • Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy .

Where you’d work

Part of the week in the office

Relocation offered

About the company

Mistral AI

  • Industry: AI

Offices in Paris, France, Amsterdam, Netherlands, London, United Kingdom, Lausanne, Switzerland

Also hiring in New York, United States, Palo Alto, United States, San Francisco, United States and 7 more places

8 of their 37 open roles are remote

Your chances

Worth a look before you spend an evening tailoring a CV for it.

  • 20 checks run
  • 4 red flags

Still hiring?

13 checks

1 red flag

How crowded?

7 checks

3 red flags

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details11 facts · Role, Location, Compensation, Employment, Company
Tech stack
  • Python
  • Kubernetes
  • PyTorch
  • LLMs
Type
Full-time
Industry
AI
Specialty
ML
Region
Europe
Pay period
Annual
Show 5 more factsShow less

Role

Category
Data & Analytics
Specialty
ML
Tech stack
  • Python
  • Kubernetes
  • PyTorch
  • LLMs

Location

Work model
Hybrid
Region
Europe
Offices
  • Paris, France
  • Amsterdam, Netherlands
  • London, United Kingdom
  • Lausanne, Switzerland
Relocation
Offered

Compensation

Salary
Salary by agreement
Pay period
Annual

Employment

Type
Full-time

Company

Industry
AI

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Data & AI Engineer Intern

    Salary by agreement

    • Hybrid · Paris
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    AI Engineer

    $80,000 - 210,000 / year

    • Office · CA, Austin
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior AI/ML Engineer

    $200,000 - 260,000 / year

    • Office · San Francisco
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Member of Technical Staff, ML Infra

    Salary by agreement

    • Office · San Francisco
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason