Junior/Senior or Staff Software Engineer, Inference / Compute Infrastructure Engineering
Together AI · India
Posted 28 days ago
or apply directly on Together AI's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 28 days Together AI's roles stay open a median of 45 days
- Reposted
- No
- Salary listed
- No 72% of Together AI's roles list one
- Ghost-job risk at Together AI
- high 42 stale, 0 reposted of 61 open
- Hiring momentum
- 77 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-17
Measured from postings appearing on and disappearing from Together AI's own greenhouse board since 2026-08-03. Full hiring picture for Together AI.
About this role
In this role, you will be responsible for building a Kubernetes-native control plane to manage a GPU inference fleet. Your tasks will include designing APIs for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. You will also collaborate with the inference/ML platform team to encode necessary cluster shapes and engineer the infrastructure code with a focus on quality and developer experience.
- benefits
- 1/5
- freshness
- 4/5
- career value
- 4/5
- role clarity
- 5/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Together AI as an employer.
What you need
- Strong software engineering background in Go, Python, Rust, or similar
- Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
- Experience building software control planes or orchestration systems that model state and reconcile it over time
- Experience with event-driven systems
- A product mindset with experience building internal platforms or APIs consumed by other engineering teams
Nice to have
- Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals
- Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
- Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
- Systems programming in Rust or Go
Worth weighing
- Remote position in India
- No specific salary or benefits mentioned
- Focus on operational and runtime complexity may require deep technical understanding
- Role involves both development and production support responsibilities
Summarised from Together AI's posting. Read the full original.
Listed by Together AI on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.
More roles at Together AI
- Strategic Finance Senior Associate - ComputeSan Francisco
- Senior Product Manager, Model APIs & Developer ExperienceSan Francisco
- Manager, International Cloud SourcingSan Francisco or NYC
- Junior/Senior/Staff Software Engineer, Inference / Compute Infrastructure EngineeringAmsterdam
- HR Coordinator- AmsterdamAmsterdam, Netherlands
- Staff Software Engineer, Inference / Compute Infrastructure EngineeringLondon & Amsterdam
- Senior Software Engineer — Infra Agent Systems Remote IndiaIndia
- Senior Software Engineer — Infra Agent Systems UKLondon