Pantheon's mission is to build and deploy billions of useful robots to automate physical labor, starting with bi-manual manipulators for industry. Industrial labor makes up almost $1T of U.S. spend, and yet over 439,000 jobs remain unfilled. Pantheon is building a future where labor is as cheap as energy and raw materials.
The approach is full-stack: owning hardware, ML, and deployment infrastructure is all required to reach 99.9%+ robotic reliability. Robotics is an engineering, operational, and infrastructure problem as much as it is a research one.
The Opportunity
This is Pantheon's first dedicated training infrastructure hire. You will own the compute the research team runs on and build most of it from scratch — GPU clusters, data clusters, scheduling, orchestration, and the training stack itself.
Pantheon is pretraining a robotics foundation model now, with runs scaling to hundreds of GPUs in the coming weeks. Your decisions about cluster architecture, job scheduling, and training reliability will directly set how fast the research team can iterate. You will inherit a live pretraining program, not a greenfield roadmap — the first 90 days are about hardening smoke-run infrastructure into runs at scale, owning the scale-up, and writing the playbook that makes large runs routine.
What You'll Do
You Should Have
Nice to Have
Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.