Toyota Research Institute (TRI) is on a mission to improve the quality of human life through innovative technology. They are seeking a Senior Research Engineer in Computer Vision to develop and deploy world foundation models for autonomous driving, collaborating closely with research scientists and multiple divisions to bridge the gap between research and production systems.
Responsibilities:
- Collaborate directly with research scientists to implement, iterate on, and evaluate new architectures, objectives, datasets, and training strategies. Translate research prototypes into clean, maintainable, reusable code that will be shared across multiple TRI teams and the broader Toyota ecosystem
- Build and maintain scalable pipelines for ingesting, converting, validating, and serving heterogeneous datasets (multi-view, multi-modal, multi-embodiment, etc.), across robotics and autonomous driving, into unified training-ready formats. Track and integrate new public and internal datasets as they become available
- Support and optimize large-scale distributed training of world foundation models on multi-GPU and multi-node clusters. Manage experiment workflows, profiling, debugging, and hyperparameter sweeps to ensure optimal performance in a timely manner
- Develop tools for dataset inspection, experiment tracking, model evaluation, GPU resource management, and visualization. Automate repetitive workflows to improve team velocity
- Work with other TRI teams and Toyota affiliates to set up shared pipelines, onboard their data, and support joint training and evaluation efforts
- Produce maintainable, well-documented code. Contribute to internal tooling and open-source releases to the scientific community