or apply directly on Nebius's site. We never take the application ourselves.
Is this posting real?
- This role has been open
- 55 days Nebius's roles stay open a median of 55 days
- Reposted
- No
- Salary listed
- No 9% of Nebius's roles list one
- Ghost-job risk at Nebius
- high 279 stale, 17 reposted of 384 open
- Hiring momentum
- 508 roles opened in the last 90 days ↑ up vs. the prior 90 days
- Last confirmed on the employer's board
- 2026-09-27
Measured from postings appearing on and disappearing from Nebius's own greenhouse board since 2026-08-03. Full hiring picture for Nebius.
About this role
The Senior Data Center Support Engineer at Nebius will be responsible for operating, troubleshooting, and maintaining large-scale bare metal GPU infrastructure that supports AI workloads. This role involves diagnosing high-priority issues across various systems, performing hardware diagnostics, and collaborating with multiple engineering teams to ensure operational recovery and incident response. The engineer will also create documentation and use automation to enhance support workflows.
- benefits
- 3/5
- freshness
- 1/5
- career value
- 4/5
- role clarity
- 4/5
- pay transparency
- 0/5
Scored from the posting itself — how clearly the role is described, how much it says about pay and benefits, and how recently it was listed. Not a judgement of Nebius as an employer.
What you need
- 5+ years of experience in data center operations, infrastructure support, systems administration, hardware support, cloud infrastructure, or similar technical environments
- Hands-on experience troubleshooting bare metal servers, Linux systems, hardware components, and networking issues in production environments
- Intermediate Linux command-line proficiency
- Experience with server hardware diagnostics, component replacement, firmware updates, BIOS configuration, driver troubleshooting, or BMC tools
- Strong networking fundamentals, including TCP/IP, VLANs, DNS, DHCP, routing/switching concepts, optics, cabling, and link-level troubleshooting
- Experience participating in incident response, escalation handling, root cause analysis, or operational recovery in high-availability environments
Nice to have
- Experience with NVIDIA GPUs, GPU servers, HPC, AI infrastructure, or high-density compute platforms
- Experience with Supermicro, Dell, HPE, Lenovo, Cisco UCS, Arista, Cisco, Juniper, or similar infrastructure platforms
- Familiarity with InfiniBand, RoCE, MPO/MTP fiber, high-speed Ethernet, or clustered compute environments
- Experience with Python, Bash, Ansible, Terraform, PowerShell, or similar automation tools
- Familiarity with Jira, Confluence, ServiceNow, Grafana, Prometheus, Kubernetes, Docker, or similar operational tools
What you get
- Competitive compensation
- Career growth and learning opportunities
- Flexibility and ownership
- Collaborative and innovative culture
- Opportunity to work on impactful AI projects
- International environment and talented teams
Worth weighing
- No specific mention of remote work options or on-site requirements
- The role involves on-call or after-hours support which may affect work-life balance
- The compensation range is broad, which may lead to variability based on negotiation or experience
Summarised from Nebius's posting. Read the full original.
Listed by Nebius on their greenhouse job board, last confirmed open on 2026-09-27. PitchMeAI is not the employer.
More roles at Nebius
- Senior Systems Software Engineer, GPU ComputeRemote - United States
- Senior Hypervisor EngineerPrague, Czech Republic; Remote - Europe
- Bare Metal Infrastructure EngineerNew Jersey, United States
- Head of AnalyticsIsrael
- Site Reliability EngineerRemote - United States
- Sr. Manager, Deal Initiation to Customer ActivationUnited States
- DC Operations Manager - Critical InfrastructureAlabama, United States
- Immigration SpecialistUnited Kingdom