Jobiglo

No results.

Member of Technical Staff, GPU Compiler

sf-tensor · San Francisco

New
285,000 - 315,000 USD/year 🇬🇧 English
LLVM MLIR CUDA ROCm PTX SASS GCN RDNA XLA TVM Triton torch.compiler C++ Rust Formal verification SMT solvers StableHLO

Job description

About the role

We are building the world’s fastest GPU compiler, breaking traditional correctness constraints to enable aggressive search‑based optimization. As a Member of Technical Staff for GPU Compiler Engineering, you will create the machine that performs the search and delivers high‑performance kernels across multiple hardware targets.

Key responsibilities

  • Design and extend MLIR dialects and passes for training and inference workloads.
  • Develop the LLVM backend below ptxas, handling instruction selection, scheduling, register allocation, and direct cubin emission for NVIDIA, AMD, TPU, and Trainium.
  • Expand search‑based compiler infrastructure with agent‑ and RL‑driven program search and formal correctness proofs.
  • Implement classic compiler optimizations tuned for large‑scale training and create hybrid codegen paths when direct MLIR lowering is impractical.
  • Own testing, benchmarking, and performance regression systems, including bit‑identical hardware models.

Required profile

  • Deep experience with compiler infrastructure such as LLVM or MLIR.
  • Strong background in GPU architecture and low‑level optimization (CUDA, ROCm, PTX/SASS, GCN/RDNA).
  • Hands‑on experience with at least one GPU ISA and familiarity with ML compiler stacks (XLA, TVM, Triton, torch.compiler).
  • Solid systems programming skills in C++ and/or Rust.
  • Proven track record of building production‑grade compiler infrastructure.

Required skills

  • LLVM
  • MLIR
  • CUDA
  • ROCm
  • PTX / SASS
  • GCN / RDNA assembly
  • XLA
  • TVM
  • Triton
  • torch.compiler
  • C++
  • Rust
  • Reinforcement Learning (RL) for program search
  • Formal verification / SMT solvers
  • StableHLO / (Stable)HLO

What we offer

  • Opportunity to work on the fastest GPU compiler with direct impact on AI compute costs.
  • Small, frontier‑scale team that has pre‑trained foundation models on thousands of GPUs.
  • Relocation assistance and an onsite office in San Francisco.
  • Competitive base salary $285,000–$315,000 plus equity and benefits.

Questions fréquentes

Le salaire proposé pour ce poste est de 285-315k USD par an. Le détail figure dans l'annonce.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Source : ats:ashby

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

Apply now →

By continuing, you accept our terms of use.

Already have an account? Login

A question about this job?

Ask it here: you will get the full job summary by e-mail, right away.

💬 Chat with us on Telegram

Published 5 days ago

Expires 1 month from now

8 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

sf-tensor

San Francisco