Operations Engineer, Fleet Reliability
CoreWeave · Plano, TX / Washington, DC / Livingston, NJ
Posted about 2 months ago · $83,000 to $110,000
or apply directly on CoreWeave's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 66 days CoreWeave's roles stay open a median of 66 days
- Reposted
- No
- Salary listed
- Yes 94% of CoreWeave's roles list one
- Ghost-job risk at CoreWeave
- high 246 stale, 8 reposted of 293 open
- Hiring momentum
- 403 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-10-08
Measured from postings appearing on and disappearing from CoreWeave's own greenhouse board since 2026-08-03. Full hiring picture for CoreWeave.
About this role
As an Operations Engineer on the Fleet Reliability team at CoreWeave, you will be responsible for the provisioning, management, and uptime of supercomputing clusters. Your daily tasks will include configuring and maintaining high-performance clusters, troubleshooting hardware and software issues, monitoring system performance, and participating in on-call rotations. Additionally, you will document processes and collaborate with team members to enhance efficiency and effectiveness in operations.
- benefits
- 5/5
- freshness
- 1/5
- career value
- 4/5
- role clarity
- 4/5
- pay transparency
- 5/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of CoreWeave as an employer.
What you need
- Strong understanding of Linux system administration and internals
- Ability to troubleshoot hardware and software issues and perform system maintenance tasks consistently and reliably
- Software development or scripting languages (bash, python, powershell, etc)
Nice to have
- 2 + years of experience troubleshooting or administering data center or on-prem infrastructure (servers, storage, network or a mix)
- Grafana, Prometheus, promsql queries or similar observability platforms
- Data center environments including server racks, HVAC systems, fiber trays
- Kubernetes administration
- HPC - administering GPU-related workloads
What you get
- Medical, dental, and vision insurance - 100% paid for by CoreWeave
- Company-paid Life Insurance
- Voluntary supplemental life insurance
- Short and long-term disability insurance
- Flexible Spending Account
- Health Savings Account
Worth weighing
- Role involves participation in on-call rotations, including after hours and weekends
- The position requires access to export controlled information, which may limit applicant eligibility
- The company is in a hyper-growth stage, which may involve a fast-paced and potentially chaotic work environment
Summarised from CoreWeave's posting. Read the full original.
Listed by CoreWeave on their greenhouse job board, last confirmed open on 2026-10-08. PitchMeAI is not the employer.
More roles at CoreWeave
- Senior Production EngineerSunnyvale, CA | Bellevue, WA
- Senior Technical Project Manager - East RegionRichmond, VA / Columbus, OH
- Staff Thermal EngineerLivingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA
- Data Center Commissioning/Quality ManagerLivingston, NJ / New York, NY / Sunnyvale, CA / San Francisco, CA / Bellevue, WA
- Sr. Technical Program Manager, Capacity DeliveryLivingston, NJ / New York, NY / Sunnyvale, CA / Bellevue, WA/San Francisco, CA
- Staff Business Systems Engineer, Data Center SystemsSunnyvale, CA
- Senior Specialist Field Engineer - NetworkingLivingston, NJ / New York, NY
- Senior Machine Learning Engineer, AI InsightsNew York, NY