The opportunity
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to…
What you'll do
Analyze and characterize internal feature representations of deep multimodal perception foundation models.
Design and implement active learning and intelligent data curation algorithms: to identify rare, safety-critical edge cases.
Explore automated data labeling and validation workflows leveraging modern Vision-Language Models (VLMs).
Train, fine-tune, and evaluate deep neural networks to improve model performance.
Validate model performance on large-scale autonomous vehicle sensor datasets and simulation environments.
Currently enrolled in a PhD program in Computer Science, Robotics, Electrical: Engineering, or a related quantitative field.
What they're looking for
- Experience programming in Python and C++.
- Practical experience training, fine-tuning, and evaluating deep learning: models for computer vision or multimodal perception (e.g., using PyTorch, JAX, or TensorFlow).
- Solid understanding of modern neural architectures (e.g., Vision Transformers, multi modal sensor encoders).
- Research experience or publications in active learning, self-supervised: representation learning, core-set selection, or multi modal foundation models (VLMs).