Nebius

Data Center Support Engineer

Nebius · New Jersey, United States

Posted about 2 months ago

or apply directly on Nebius's site. We never take the application ourselves.

Is this posting real?

This role has been open
55 days
Nebius's roles stay open a median of 55 days
Reposted
No
Salary listed
No
9% of Nebius's roles list one
Ghost-job risk at Nebius
high
279 stale, 17 reposted of 384 open
Hiring momentum
508 roles opened in the last 90 days
↑ up vs. the prior 90 days
Last confirmed on the employer's board
2026-09-27

Measured from postings appearing on and disappearing from Nebius's own greenhouse board since 2026-08-03. Full hiring picture for Nebius.

About this role

The Senior Data Center Support Engineer at Nebius will be responsible for operating, troubleshooting, and maintaining large-scale bare metal GPU infrastructure that supports AI workloads. This role involves diagnosing high-priority issues across various systems, performing hardware diagnostics, and collaborating with multiple engineering teams to ensure operational recovery and incident response. The engineer will also create documentation and use automation to enhance support workflows.

Our read on this posting2.4out of 5
benefits
3/5
freshness
1/5
career value
4/5
role clarity
4/5
pay transparency
0/5

Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Nebius as an employer.

What you need

  • 5+ years of experience in data center operations, infrastructure support, systems administration, hardware support, cloud infrastructure, or similar technical environments
  • Hands-on experience troubleshooting bare metal servers, Linux systems, hardware components, and networking issues in production environments
  • Intermediate Linux command-line proficiency
  • Experience with server hardware diagnostics, component replacement, firmware updates, BIOS configuration, driver troubleshooting, or BMC tools
  • Strong networking fundamentals, including TCP/IP, VLANs, DNS, DHCP, routing/switching concepts, optics, cabling, and link-level troubleshooting
  • Experience participating in incident response, escalation handling, root cause analysis, or operational recovery in high-availability environments

Nice to have

  • Experience with NVIDIA GPUs, GPU servers, HPC, AI infrastructure, or high-density compute platforms
  • Experience with Supermicro, Dell, HPE, Lenovo, Cisco UCS, Arista, Cisco, Juniper, or similar infrastructure platforms
  • Familiarity with InfiniBand, RoCE, MPO/MTP fiber, high-speed Ethernet, or clustered compute environments
  • Experience with Python, Bash, Ansible, Terraform, PowerShell, or similar automation tools
  • Familiarity with Jira, Confluence, ServiceNow, Grafana, Prometheus, Kubernetes, Docker, or similar operational tools

What you get

  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and ownership
  • Collaborative and innovative culture
  • Opportunity to work on impactful AI projects
  • International environment and talented teams

Worth weighing

  • No specific mention of remote work options or on-site requirements
  • The role involves on-call or after-hours support which may affect work-life balance
  • The compensation range is broad, which may lead to variability based on negotiation or experience

Summarised from Nebius's posting. Read the full original.

Listed by Nebius on their greenhouse job board, last confirmed open on 2026-09-27. PitchMeAI is not the employer.

More roles at Nebius

All 384 open roles at Nebius →