Junior Technical Program Manager — Infrastructure Operations
Togetherai
San Francisco$150k – $175k22d ago
Looking for more like this? See all Project Manager jobs.
About the role
About the Role
Together AI runs one of the most demanding GPU fleets in the industry. Keeping that fleet healthy - every node online, every GPU performing, every datacenter transition running on schedule - is operationally complex and genuinely high-stakes. We're looking for a Junior TPM to own that operational reality.
This is not a coordination or status-reporting role. You will own the end-to-end node lifecycle - from the moment a node goes down through repair, return, and re-integration - and you'll drive the cross-functional work to close every gap as fast as possible. You'll manage datacenter bring-ups, hunt down GPU utilization loss, and build the processes and dashboards that make our fleet operations more visible and accountable over time.
The environment moves fast and doesn't always come with a clear playbook. Much of what you'll work on is genuinely novel - y
More at Togetherai
- Technical Account Manager (TAM), GPU ClusterSan Francisco
- Staff Engineer, Distributed Storage and HPC & AI InfrastructureSan Francisco · $250k – $300k
- Manager, Infrastructure Strategy & OperationsSan Francisco
- Customer Support Engineer (Inference)San Francisco, CA · $160k – $230000k
- Senior Technical Recruiter, AI/ML ResearchSan Francisco · $165k – $210k
- Engineering Manager, Site Reliability EngineeringSan Francisco