PitchMeAI
Altera Digital Health APAC

Expert Site Reliability Engineer

Altera Digital Health APAC · Illinois, United States

  • On site
  • Full-time
  • $102,500 / year
  • Illinois, United States
Tailored resumekeyword-matched to this role.
Hiring managerwe find who's hiring.
Intro emaildrafted to reach them directly.

Job highlights

  • Ensure healthcare platform reliability and scalability.
  • Lead incident response and root cause analysis.
  • Automate operations using scripting and IaC.
  • Collaborate with engineering and cloud teams.
  • Impact patient lives by improving healthcare delivery.

About the role

Site Reliability Engineer (SRE) - Remote

Overview

As a Site Reliability Engineer (SRE) at Altera, you will be responsible for ensuring the reliability, scalability, and performance of our hosted healthcare platforms. This role blends software and systems engineering to enhance service availability, automate operations, and improve the customer experience. You will act as a technical leader in monitoring, troubleshooting, incident response, and continuous improvement across our cloud and hybrid environments.

Key Responsibilities

  • Maintain and improve the reliability, availability, and performance of our production environments.
  • Lead the investigation and resolution of complex application, database, and infrastructure issues.
  • Participate in incident management, conduct root cause analysis (RCA), and contribute to post-incident reviews to prevent future occurrences.
  • Define and measure Service Level Indicators (SLIs) and Objectives (SLOs) to meet our service commitments.
  • Develop proactive monitoring and alerting strategies to identify and resolve issues before they impact customers.
  • Automate operational tasks using scripting and Infrastructure-as-Code (IaC) to improve efficiency.
  • Partner with engineering and cloud teams to refine deployment, monitoring, and support processes.
  • Provide technical leadership during major incidents and act as a key escalation point for critical issues.

Qualifications

Experience:

  • 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
  • Monitoring & Observability: Strong experience with APM tools such as LogicMonitor, AppDynamics, Azure Monitor, SentryOne, Dynatrace, Datadog, or New Relic.
  • Microsoft Stack: Deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
  • Database Skills: Strong SQL Server experience, including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
  • Cloud & Networking: Experience with Azure cloud environments and a solid understanding of networking fundamentals (DNS, TCP/IP, load balancing, firewalls).
  • ITSM & ITIL: Familiarity with ServiceNow (or other ITSM platforms) and ITIL principles.

Preferred Skills

  • Scripting with PowerShell, Python, or similar languages.
  • Infrastructure as Code (Terraform, ARM Templates, Bicep).
  • CI/CD pipelines and deployment automation (Azure DevOps, GitHub Actions).
  • Experience with Kubernetes and containerized workloads.
  • Experience implementing SLOs, SLIs, and Error Budgets.
  • Experience in a healthcare technology or patient care environment.

Education

Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred; equivalent professional experience will be considered.

Working Arrangements

  • This is a remote position open to candidates within the United States.
  • You will participate in an on-call rotation to support our 24x7 healthcare environment.
  • Occasional after-hours work is required for activations, upgrades, and major incidents.

Travel

Travel is not a requirement for this role.

Salary Range

$95,000 - $110,000

Why Altera?

At Altera Digital Health, you will have the opportunity to profoundly impact the lives of patients by empowering healthcare providers to deliver superior care. You will join a passionate and gifted team committed to innovation and excellence. We offer a competitive compensation and benefits package and the opportunity to work in a fast-paced and dynamic environment.

Key skills/competency

  • Site Reliability Engineer
  • Cloud Computing (Azure)
  • Monitoring and Observability
  • Infrastructure as Code
  • Windows Server Administration
  • SQL Server
  • Networking
  • Incident Management
  • Automation
  • ITSM/ITIL

Skills & topics

  • Site Reliability Engineer
  • SRE
  • Cloud Engineer
  • DevOps Engineer
  • Systems Engineer
  • Azure
  • SQL Server
  • Windows Server
  • Monitoring
  • Automation
  • Infrastructure as Code
  • ITSM
  • ITIL
  • Healthcare Technology
  • Remote

How to get hired

  • Tailor your resume: Highlight 7+ years of experience in enterprise applications, cloud environments, and specific tools like Azure Monitor, Datadog, SQL Server, and ServiceNow.
  • Showcase technical skills: Emphasize your proficiency in Windows Server, .NET, SQL Server performance tuning, Azure cloud, and networking fundamentals.
  • Quantify achievements: Use metrics to demonstrate your impact on reliability, performance, and automation initiatives.
  • Prepare for technical interviews: Be ready to discuss complex troubleshooting scenarios, incident management processes, and your experience with IaC and scripting.
  • Highlight relevant experience: If you have experience in healthcare technology or with Kubernetes, be sure to mention it.

Technical preparation

Master Windows Server and SQL Server performance tuning.,Practice Azure cloud and networking fundamentals.,Build automation scripts with PowerShell or Python.,Familiarize with IaC tools like Terraform.

Behavioral questions

Describe a major incident you resolved.,How do you define and measure reliability?,How do you collaborate with development teams?,How do you prioritize urgent issues?

Frequently asked questions

What is the salary range for the Site Reliability Engineer role at Altera Digital Health APAC?
The salary range for the Site Reliability Engineer position at Altera Digital Health APAC is between $95,000 and $110,000 annually. This range is determined by factors such as internal equity, market data, and the applicant's skills and experience.
Is the Site Reliability Engineer position remote, and what are the location requirements?
Yes, the Site Reliability Engineer position is a remote role and is open to candidates within the United States. You will not be required to travel for this position.
What specific Microsoft technologies are important for this Site Reliability Engineer role?
For this Site Reliability Engineer role, deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon is crucial. This expertise is essential for managing and supporting the production environments.
What kind of experience is needed to be considered for the Site Reliability Engineer role?
To be considered for the Site Reliability Engineer role, you need at least 7 years of experience supporting enterprise applications, infrastructure, or cloud environments. Strong experience with monitoring tools, Microsoft stack, SQL Server, and Azure cloud environments is also required.
Does Altera Digital Health APAC require on-call for the Site Reliability Engineer position?
Yes, as a Site Reliability Engineer at Altera Digital Health APAC, you will participate in an on-call rotation to ensure 24x7 support for their healthcare environment. Occasional after-hours work may also be required.
What are the educational requirements for the Site Reliability Engineer role?
A Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred for the Site Reliability Engineer role. However, equivalent professional experience will also be considered.
What are the key responsibilities of a Site Reliability Engineer at Altera Digital Health APAC?
Key responsibilities include maintaining and improving the reliability of production environments, leading incident investigations and root cause analysis, developing proactive monitoring and alerting strategies, and automating operational tasks using scripting and Infrastructure-as-Code.
What preferred skills would make a candidate stand out for the Site Reliability Engineer role?
Preferred skills for the Site Reliability Engineer role include scripting with PowerShell or Python, experience with Infrastructure as Code (Terraform, ARM Templates, Bicep), CI/CD pipelines, Kubernetes, implementing SLOs/SLIs, and experience in healthcare technology.