Staff Software Engineer, Inference / Compute Infrastructure Engineering
Together AI · London & Amsterdam
Posted 27 days ago
or apply directly on Together AI's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 27 days Together AI's roles stay open a median of 45 days
- Reposted
- No
- Salary listed
- No 72% of Together AI's roles list one
- Ghost-job risk at Together AI
- high 42 stale, 0 reposted of 61 open
- Hiring momentum
- 77 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-17
Measured from postings appearing on and disappearing from Together AI's own greenhouse board since 2026-08-03. Full hiring picture for Together AI.
About this role
The Staff Software Engineer will focus on building a Kubernetes-native control plane for managing a GPU inference fleet, enabling the inference team to efficiently provision and manage resources via a self-service API. Responsibilities include designing lifecycle management software, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. The role emphasizes a product mindset, requiring ownership of both software delivery and operational support in production environments.
- benefits
- 1/5
- freshness
- 4/5
- career value
- 5/5
- role clarity
- 5/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Together AI as an employer.
What you need
- Strong software engineering background in Go, Python, Rust, or similar
- Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
- Experience building software control planes or orchestration systems that model state and reconcile it over time
- Experience with event-driven systems designing and building software around message queues, event streams, or pub/sub
- A product mindset with experience building internal platforms or APIs consumed by other engineering teams
Nice to have
- Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design)
- Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
- Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
- Systems programming in Rust or Go
Worth weighing
- No salary or benefits listed
- Role involves significant operational responsibility alongside software development
- Focus on GPU infrastructure may require specialized knowledge
- Emphasis on self-service and automation implies a fast-paced, potentially high-pressure environment
Summarised from Together AI's posting. Read the full original.
Listed by Together AI on their greenhouse job board, last confirmed open on 2026-09-17. PitchMeAI is not the employer.
More roles at Together AI
- Strategic Finance Senior Associate - ComputeSan Francisco
- Senior Product Manager, Model APIs & Developer ExperienceSan Francisco
- Manager, International Cloud SourcingSan Francisco or NYC
- Junior/Senior/Staff Software Engineer, Inference / Compute Infrastructure EngineeringAmsterdam
- Junior/Senior or Staff Software Engineer, Inference / Compute Infrastructure EngineeringIndia
- HR Coordinator- AmsterdamAmsterdam, Netherlands
- Senior Software Engineer — Infra Agent Systems Remote IndiaIndia
- Senior Software Engineer — Infra Agent Systems UKLondon