Staff Software Engineer, Inference / Compute Infrastructure Engineering
Together AI · San Francisco
Posted about 2 months ago · $240,000 - $280,000
or apply directly on Together AI's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 53 days Together AI's roles stay open a median of 54 days
- Reposted
- No
- Salary listed
- Yes 72% of Together AI's roles list one
- Ghost-job risk at Together AI
- high 45 stale, 0 reposted of 61 open
- Hiring momentum
- 77 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-26
Measured from postings appearing on and disappearing from Together AI's own greenhouse board since 2026-08-03. Full hiring picture for Together AI.
About this role
The Staff Software Engineer will focus on building a Kubernetes-native control plane for managing GPU inference fleets. Responsibilities include designing a self-service API for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. The role emphasizes a product mindset, requiring the engineer to not only develop software but also operate and support it in production environments.
- benefits
- 3/5
- freshness
- 1/5
- career value
- 4/5
- role clarity
- 5/5
- pay transparency
- 5/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Together AI as an employer.
What you need
- Strong software engineering background in Go, Python, Rust, or similar
- Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
- Experience building software control planes or orchestration systems that model state and reconcile it over time
- Experience with event-driven systems designing and building software around message queues, event streams, or pub/sub
- A product mindset with experience building internal platforms or APIs consumed by other engineering teams
Nice to have
- Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design)
- Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
- Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
- Systems programming in Rust or Go
What you get
- Competitive compensation
- Startup equity
- Health insurance
- Other competitive benefits
- US base salary range of $240,000 - $280,000 + equity + benefits
Worth weighing
- No specific mention of remote work options
- Focus on GPU infrastructure may limit exposure to other tech stacks
- Role requires both development and operational responsibilities, which may lead to a heavier workload
Summarised from Together AI's posting. Read the full original.
Listed by Together AI on their greenhouse job board, last confirmed open on 2026-09-26. PitchMeAI is not the employer.
More roles at Together AI
- Senior Software Engineer Together Cloud InfrastructureAmsterdam
- Senior Program Manager, Data Center DeliveryRemote
- Associate, Infrastructure Strategy & OperationsSan Francisco
- Senior Software Engineer - Together Cloud InfrastructureSan Francisco
- Senior Technical Recruiter, AI/ML ResearchSan Francisco
- AI Infrastructure Systems EngineerSan Francisco
- Data Center Operations CoordinatorSan Francisco
- Technical Support Engineer (Inference) - India WeekendsPune or Bangalore, India