Jobiglo

No results.

Distillation Lead – Model Compression for Autonomous AI

waabi · San Francisco

New
🇬🇧 English
Python PyTorch JAX TensorRT ONNX CUDA distributed training

Job description

About the role

Waabi is a leader in Physical AI, building the technology behind autonomous trucks and robotaxis. As the Distillation Lead you will own the strategy and execution for model distillation across Waabi’s AI stack, ensuring high‑performing models run efficiently in both onboard vehicles and large‑scale simulation pipelines.

Key responsibilities

  • Define and drive the technical strategy for model distillation and compression across perception, world models, and planning.
  • Design, implement, and scale state‑of‑the‑art distillation pipelines, including generative model distillation, quantization‑aware training, knowledge distillation, pruning, low‑rank factorization, and speculative decoding.
  • Collaborate with ML Platform, Infrastructure, Onboard Autonomy, and Simulation teams to integrate compressed models into production, meeting latency, memory, and throughput targets.
  • Define rigorous benchmarks and evaluation frameworks to assess efficiency vs. quality trade‑offs across hardware targets.
  • Mentor researchers and engineers, champion best practices, and contribute to publications and open‑source projects.

Required profile

  • Extensive hands‑on experience with model distillation, quantization, pruning, and compression for large‑scale neural networks in production.
  • Strong research and engineering foundation; Bachelor’s or Master’s in ML, Computer Vision, Robotics, or equivalent industry experience.
  • Proven technical leadership, setting direction and driving projects from concept to deployment.
  • Experience collaborating with infrastructure, platform, and autonomy teams under real engineering constraints.
  • Excellent communication skills to convey complex trade‑offs to diverse audiences.

Required skills

  • Python
  • PyTorch (or JAX)
  • Distributed training at large scale
  • TensorRT, ONNX, custom CUDA kernels
  • Model distillation, quantization‑aware training, pruning, low‑rank factorization

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec waabi.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Source : ats:lever

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

A question about this job?

Ask it here: you will get the full job summary by e-mail, right away.

💬 Chat with us on Telegram

Published 1 day ago

Expires 1 month from now

1 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

waabi

San Francisco