Together AI

Junior/Senior or Staff Software Engineer, Inference / Compute Infrastructure Engineering

Together AI · India

Posted 28 days ago

or apply directly on Together AI's site. We never take the application ourselves.

Is this posting real?

This role has been open
28 days
Together AI's roles stay open a median of 45 days
Reposted
No
Salary listed
No
72% of Together AI's roles list one
Ghost-job risk at Together AI
high
42 stale, 0 reposted of 61 open
Hiring momentum
77 roles opened in the last 90 days
↑ up vs. the prior 90 days
Last confirmed on the employer's board
2026-09-17

Measured from postings appearing on and disappearing from Together AI's own greenhouse board since 2026-08-03. Full hiring picture for Together AI.

About this role

In this role, you will be responsible for building a Kubernetes-native control plane to manage a GPU inference fleet. Your tasks will include designing APIs for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. You will also collaborate with the inference/ML platform team to encode necessary cluster shapes and engineer the infrastructure code with a focus on quality and developer experience.

Our read on this posting2.8out of 5
benefits
1/5
freshness
4/5
career value
4/5
role clarity
5/5
pay transparency
0/5

Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Together AI as an employer.

What you need

  • Strong software engineering background in Go, Python, Rust, or similar
  • Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
  • Experience building software control planes or orchestration systems that model state and reconcile it over time
  • Experience with event-driven systems
  • A product mindset with experience building internal platforms or APIs consumed by other engineering teams

Nice to have

  • Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals
  • Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
  • Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
  • Systems programming in Rust or Go

Worth weighing

  • Remote position in India
  • No specific salary or benefits mentioned
  • Focus on operational and runtime complexity may require deep technical understanding
  • Role involves both development and production support responsibilities

Summarised from Together AI's posting. Read the full original.

Listed by Together AI on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.

More roles at Together AI

All 61 open roles at Together AI