Together AI runs one of the most demanding GPU fleets in the industry. Keeping that fleet healthy — every node online, every GPU performing, every datacenter transition running on schedule — is operationally complex and genuinely high-stakes. We're looking for a Junior TPM to own that operational reality.
This is not a coordination or status-reporting role. You will own the end-to-end node lifecycle — from the moment a node goes down through repair, return, and re-integration — and you'll drive the cross-functional work to close every gap as fast as possible. You'll manage datacenter bring-ups, hunt down GPU utilization loss, and build the processes and dashboards that make our fleet operations more visible and accountable over time.
The environment moves fast and doesn't always come with a clear playbook. Much of what you'll work on is genuinely novel — you'll be figuring things out alongside engineers who are building at the frontier. If that sounds like an obstacle, this isn't the right role. If it sounds like the best possible way to learn, keep reading.
Together AI is a research-driven artificial intelligence company. We believe open and transparent AI systems will drive innovation and create the best outcomes for society, and together we are on a mission to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models. We have contributed to leading open-source research, models, and datasets to advance the frontier of AI, and our team has been behind technological advancement such as FlashAttention, Hyena, FlexGen, and RedPajama. We invite you to join a passionate group of researchers in our journey in building the next generation AI infrastructure.
We offer competitive compensation, startup equity, health insurance and other competitive benefits. The US base salary range for this full-time position is: $150,000 - $175,000 + equity + benefits. Our salary ranges are determined by location, level and role. Individual compensation will be determined by experience, skills, and job-related knowledge. This is a hybrid role based in the Bay Area.
Together AI is an Equal Opportunity Employer and is proud to offer equal employment opportunity to everyone regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, and more.
Please see our privacy policy at https://www.together.ai/privacy
Open-source AI cloud. Fast inference and fine-tuning for open models.
View company profileYou'll be redirected to the company's application page
Get roles like this daily
Join our Telegram channels for curated job alerts
Hey! Looking for your next role in Web3, AI, or Robotics? I can help.
Sign up to save jobs and access them across all your devices.