You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Early Window: Be the first to open itNo views yetCloses in
Company hidden

Research Scientist / Engineer

  • Remote

Salary

Not stated

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

You'll build the distributed systems that train Luma's large-scale multimodal models across thousands of GPUs, so researchers can focus on innovation on top of reliable, efficient, scalable infrastructure.

This is hard PyTorch, CUDA, and distributed-systems work — advanced parallelism, training stability, and utilization across massive clusters. It fits an engineer who's solved real problems training foundation models at scale. If you haven't worked at the level of FSDP and multi-node training, this is the wrong depth.

What You'll Own

Design, implement, and optimize efficient distributed training systems for models across thousands of GPUs.

Research and implement advanced parallelization (FSDP, Tensor Parallel, Pipeline Parallel, Expert Parallel).

Build monitoring, visualization, and debugging tools for large-scale training runs.

Optimize training stability, convergence, and resource utilization across massive clusters.

First 90 Days

One way the first 90 could unfold.

Days 1–30 — Immerse & Diagnose: Learn the current training stack and where stability and utilization hurt at scale.

Days 30–60 — Ship & Validate: Land a parallelization or stability improvement that measurably helps a real training run.

Days 60–90 — Scale & Systemize: Build the monitoring and tooling that keeps large runs reliable and efficient.

What You Bring

Extensive distributed PyTorch training and parallelisms in foundation-model training.

Deep understanding of GPU clusters, networking, and storage systems.

Familiarity with communication libraries (NCCL, MPI) and distributed-system optimization.

Requirements

  • Strong Linux systems administration and scripting.
  • Experience managing training runs across 100+ GPUs.
  • Experience with containerization, orchestration, and cloud infrastructure.
  • About Luma: Luma's mission is to build unified general intelligence that can generate, understand, and operate in the physical world. We believe multimodality is critical for intelligence — the next step beyond language models comes from vision. Luma is an equal opportunity employer.

Where you’d work

Fully remote

You can work from

  • Europe

About the company

Company hidden

  • Industry: AI

Your chances

Still hiring, not crowded yet, and you'd be among the first.

  • 16 checks run
  • 5 good signs
  • 0 red flags

Still hiring?

11 checks

Actively hiring

In its favour3

  • Still on the company's own careers site, checked 1 h agoModerate evidence
  • Found in the last 48 hours, before the big job boardsModerate evidence
  • The company opened 5 roles and closed 3 in the last 2 weeks: hiring is movingModerate evidence

How crowded?

5 checks

Low

In its favour2

  • In its Early Window: not on the big job boards yetStrong evidence
  • Senior level: far fewer people qualifySlight evidence

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details10 facts · Role, Location, Compensation, Employment, Company
Tech stack
  • Linux
  • PyTorch
  • CUDA
Type
Full-time
Industry
AI
Specialty
Data Science
Region
Europe
Pay period
Annual
Show 4 more factsShow less

Role

Category
Data & Analytics
Specialty
Data Science
Tech stack
  • Linux
  • PyTorch
  • CUDA

Location

Work model
Remote
Region
Europe
Remote from
  • Europe

Compensation

Salary
Salary by agreement
Pay period
Annual

Employment

Type
Full-time

Company

Industry
AI

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Copy of Research Scientist / Engineer

    Salary by agreement

    • Hybrid · London
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Research Scientist / Engineer

    Salary by agreement

    • Remote · Europe
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Copy of Research Scientist / Engineer

    Salary by agreement

    • Hybrid · London
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Data Science Lead

    Salary by agreement

    • Hybrid · London
    • Lead & Manager

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason