Together AI

Staff Software Engineer, Inference / Compute Infrastructure Engineering

Together AI · San Francisco

Posted about 2 months ago · $240,000 - $280,000

or apply directly on Together AI's site. We never take the application ourselves.

Is this posting real?

This role has been open
53 days
Together AI's roles stay open a median of 54 days
Reposted
No
Salary listed
Yes
72% of Together AI's roles list one
Ghost-job risk at Together AI
high
45 stale, 0 reposted of 61 open
Hiring momentum
77 roles opened in the last 90 days
↑ up vs. the prior 90 days
Last confirmed on the employer's board
2026-09-26

Measured from postings appearing on and disappearing from Together AI's own greenhouse board since 2026-08-03. Full hiring picture for Together AI.

About this role

The Staff Software Engineer will focus on building a Kubernetes-native control plane for managing GPU inference fleets. Responsibilities include designing a self-service API for the inference team, automating self-healing processes, and ensuring the reliability of the provisioning pipeline. The role emphasizes a product mindset, requiring the engineer to not only develop software but also operate and support it in production environments.

Our read on this posting3.6out of 5
benefits
3/5
freshness
1/5
career value
4/5
role clarity
5/5
pay transparency
5/5

Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Together AI as an employer.

What you need

  • Strong software engineering background in Go, Python, Rust, or similar
  • Experience with durable workflow orchestration tools such as Temporal, Cadence, or equivalent
  • Experience building software control planes or orchestration systems that model state and reconcile it over time
  • Experience with event-driven systems designing and building software around message queues, event streams, or pub/sub
  • A product mindset with experience building internal platforms or APIs consumed by other engineering teams

Nice to have

  • Exposure to bare-metal provisioning (PXE/iPXE, Redfish/IPMI, BMC) and/or networking fundamentals (VLANs, BGP, fabric design)
  • Experience with GPU cluster software stacks (NCCL, CUDA, InfiniBand/RoCE)
  • Prior work at a hyperscaler, GPU cloud, or datacenter-scale infrastructure organization
  • Systems programming in Rust or Go

What you get

  • Competitive compensation
  • Startup equity
  • Health insurance
  • Other competitive benefits
  • US base salary range of $240,000 - $280,000 + equity + benefits

Worth weighing

  • No specific mention of remote work options
  • Focus on GPU infrastructure may limit exposure to other tech stacks
  • Role requires both development and operational responsibilities, which may lead to a heavier workload

Summarised from Together AI's posting. Read the full original.

Listed by Together AI on their greenhouse job board, last confirmed open on 2026-09-26. PitchMeAI is not the employer.

More roles at Together AI

All 61 open roles at Together AI →