
Senior HPC Software Engineer
Ford Motor Company · United States
- Hybrid
- Full-time
- $192,000 / year
- United States
Job highlights
- Modernize and scale on-premise HPC platform.
- Administer RHEL systems and troubleshoot complex issues.
- Develop tools, scripts, and automation with Python/Go.
- Enhance reliability, usability, and operational maturity.
- Collaborate with teams on demanding engineering/AI workloads.
About the role
About the Role
We are seeking a senior technical contributor to help support, modernize, and scale our on premise high performance computing platform. This role will work across Linux systems administration, HPC operations, Kubernetes-based services, automation, observability, software tooling, and user-facing platform delivery. The ideal candidate has deep experience administering RHEL based systems in complex compute environments and is comfortable troubleshooting issues across operating systems, schedulers, storage, networking, containers, applications, and user workloads.
This person will play a key role in improving the reliability, usability, and operational maturity of the platform. They will help develop and maintain core HPC services, support users running demanding engineering and AI/ML workloads, and create tooling, scripts, APIs, and integrations. Strong software engineering fundamentals are important, including experience with Python, Go, or similar languages, Git-based development workflows, code reviews, testing practices, CI/CD pipelines, documentation, and maintainable code design. Experience with Slurm or other workload managers is highly valued.
We are looking for someone who can balance strong technical depth with a user-focused delivery mindset. This role requires the ability to work collaboratively with platform engineers, application teams, and technical users to identify pain points, resolve production issues, document repeatable processes, and build durable improvements. The right candidate will be pragmatic, a team player, comfortable in a fast-moving environment, and motivated by making complex, massive on-prem infrastructure easier to operate, automate, observe, and continuously improve.
Key Responsibilities
- Administer, troubleshoot, and improve RHEL based high performance computing environments supporting CPU and GPU workloads.
- Create and maintain HPC services across compute, storage, networking, scheduling, Kubernetes, and observability.
- Develop tools, scripts, APIs, integrations, and automation using Python, Go, Bash, or similar languages.
- Apply software engineering best practices, including Git workflows, code reviews, testing, modular design, and CI/CD.
- Support and help update HPC scheduling environments, with Slurm experience preferred.
- Improve monitoring, alerting, dashboards, and operational visibility using Grafana, Prometheus, Dynatrace, and related tools.
- Partner with users, customers, and internal engineering teams to understand requirements, resolve issues, and improve platform usability.
- Create and maintain documentation, architecture notes, user guides, and operational procedures.
- Drive platform modernization focused on reliability, scalability, automation, security, and maintainability.
Qualifications
- Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent experience
- 10+ years of experience in systems engineering, infrastructure engineering, platform engineering, or a related technical role.
- Strong Linux systems administration experience, preferably with RHEL.
- Experience with Slurm, PBS, or another HPC workload manager.
- Experience creating APIs, applications, and services that support platform operations and user workflows.
- Experience supporting production compute, infrastructure, and large-scale technical environments.
- Hands-on experience with scripting and software development using Python, Go, Bash, or similar languages.
- Familiarity with CI/CD concepts, GitHub, and modern software delivery practices.
- Strong troubleshooting skills across operating systems, services, networking, storage, and application layers.
- Ability to write clear documentation and communicate effectively with both technical and non-technical stakeholders.
- Strong ownership mindset with the ability to drive issues to resolution.
- Ability to use independent judgement to make sound technical decisions.
You may not check every box, or your experience may look a little different from what we've outlined, but if you think you can bring value to Ford Motor Company, we encourage you to apply!
Benefits and Perks
- Immediate medical, dental, and prescription drug coverage
- Flexible family care, parental leave, new parent ramp-up programs, subsidized back-up child care and more
- Vehicle discount program for employees and family members, and management leases
- Tuition assistance
- Established and active employee resource groups
- Paid time off for individual and team community service
- A generous schedule of paid holidays, including the week between Christmas and New Year’s Day
- Paid time off and the option to purchase additional vacation time.
Additional Information
- This position is a salary grade 8.
- Salary range: $113,580-$192,900.
- Visa Sponsorship is not provided for this role.
- Candidates for positions with Ford Motor Company must be legally authorized to work in the United States. Verification of employment eligibility will be required at the time of hire.
- We are an Equal Opportunity Employer committed to a culturally diverse workforce.
Key skills/competency
- Senior HPC Software Engineer
- Linux Systems Administration
- HPC Operations
- Kubernetes
- Automation
- Observability
- Software Tooling
- Python
- Go
- Slurm
Skills & topics
- Senior HPC Software Engineer
- HPC
- High Performance Computing
- Linux
- RHEL
- Systems Administration
- Kubernetes
- Python
- Go
- Slurm
- Automation
- Observability
- Platform Engineering
- Infrastructure Engineering
- AI/ML Workloads
- CI/CD
- Software Engineering
How to get hired
- Tailor your resume: Highlight RHEL administration, HPC experience, Python/Go scripting, and CI/CD practices.
- Showcase software engineering skills: Emphasize Git workflows, code reviews, testing, and modular design.
- Demonstrate problem-solving: Provide examples of troubleshooting complex issues across systems, schedulers, and storage.
- Align with company values: Express a user-focused delivery mindset and collaborative team spirit.
- Prepare for technical interviews: Expect questions on Linux, HPC, Kubernetes, and scripting languages.
Technical preparation
Behavioral questions
Frequently asked questions
- What are the key technical skills required for the Senior HPC Software Engineer role at Ford?
- The Senior HPC Software Engineer role at Ford requires strong Linux systems administration (preferably RHEL), experience with HPC workload managers like Slurm, proficiency in scripting and software development with Python or Go, and familiarity with Kubernetes, automation, and observability tools. A solid understanding of software engineering best practices, including Git workflows and CI/CD, is also essential.
- Does Ford offer visa sponsorship for the Senior HPC Software Engineer position?
- No, visa sponsorship is not provided for this Senior HPC Software Engineer role at Ford Motor Company. Candidates must be legally authorized to work in the United States.
- What is the salary range for a Senior HPC Software Engineer at Ford?
- The salary range for this Senior HPC Software Engineer position at Ford is between $113,580 and $192,900 annually. This is a salary grade 8 position.
- What is the typical work arrangement for this Senior HPC Software Engineer role?
- Based on the description focusing on 'on premise high performance computing platform', this Senior HPC Software Engineer role is likely an on-site position. While not explicitly stated, working with large-scale infrastructure typically requires physical presence.
- What kind of workloads will I support as a Senior HPC Software Engineer at Ford?
- As a Senior HPC Software Engineer at Ford, you will support demanding engineering and AI/ML workloads. This includes managing and troubleshooting complex CPU and GPU environments within the high-performance computing platform.
- What are the educational and experience requirements for this role?
- A Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent experience, is required. Additionally, 10+ years of experience in systems engineering, infrastructure engineering, or platform engineering is expected for this Senior HPC Software Engineer position.
- How important is experience with Slurm for this role?
- Experience with Slurm or other HPC workload managers is highly valued for this Senior HPC Software Engineer role. While not strictly mandatory, it is listed as a preferred qualification and will be a significant advantage for candidates.
- What does Ford offer in terms of benefits for this position?
- Ford offers comprehensive benefits including immediate medical, dental, and prescription drug coverage, flexible family care options, parental leave, a vehicle discount program, tuition assistance, employee resource groups, paid time off for community service, generous holidays, and the option to purchase additional vacation time.