PitchMeAI
Altera Digital Health APAC

Expert Site Reliability Engineer

Altera Digital Health APAC · Arizona, United States

  • On site
  • Full-time
  • $110,000 / year
  • Arizona, United States
Tailored resumekeyword-matched to this role.
Hiring managerwe find who's hiring.
Intro emaildrafted to reach them directly.

Job highlights

  • Ensure healthcare platform reliability and performance.
  • Lead incident response and root cause analysis.
  • Automate operations with scripting and IaC.
  • Define and measure SLIs and SLOs.
  • Remote role with a competitive salary.

About the role

Site Reliability Engineer (SRE) - Remote

Overview

As a Site Reliability Engineer (SRE) at Altera, you will be responsible for ensuring the reliability, scalability, and performance of our hosted healthcare platforms. This role blends software and systems engineering to enhance service availability, automate operations, and improve the customer experience. You will act as a technical leader in monitoring, troubleshooting, incident response, and continuous improvement across our cloud and hybrid environments.

Key Responsibilities

  • Maintain and improve the reliability, availability, and performance of our production environments.
  • Lead the investigation and resolution of complex application, database, and infrastructure issues.
  • Participate in incident management, conduct root cause analysis (RCA), and contribute to post-incident reviews to prevent future occurrences.
  • Define and measure Service Level Indicators (SLIs) and Objectives (SLOs) to meet our service commitments.
  • Develop proactive monitoring and alerting strategies to identify and resolve issues before they impact customers.
  • Automate operational tasks using scripting and Infrastructure-as-Code (IaC) to improve efficiency.
  • Partner with engineering and cloud teams to refine deployment, monitoring, and support processes.
  • Provide technical leadership during major incidents and act as a key escalation point for critical issues.

Qualifications

Experience
  • 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
  • Monitoring & Observability: Strong experience with APM tools such as LogicMonitor, AppDynamics, Azure Monitor, SentryOne, Dynatrace, Datadog, or New Relic.
  • Microsoft Stack: Deep knowledge of Windows Server administration, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
  • Database Skills: Strong SQL Server experience, including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
  • Cloud & Networking: Experience with Azure cloud environments and a solid understanding of networking fundamentals (DNS, TCP/IP, load balancing, firewalls).
  • ITSM & ITIL: Familiarity with ServiceNow (or other ITSM platforms) and ITIL principles.
Preferred Skills
  • Scripting with PowerShell, Python, or similar languages.
  • Infrastructure as Code (Terraform, ARM Templates, Bicep).
  • CI/CD pipelines and deployment automation (Azure DevOps, GitHub Actions).
  • Experience with Kubernetes and containerized workloads.
  • Experience implementing SLOs, SLIs, and Error Budgets.
  • Experience in a healthcare technology or patient care environment.
Education

Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred; equivalent professional experience will be considered.

Working Arrangements

This is a remote position open to candidates within the United States. You will participate in an on-call rotation to support our 24x7 healthcare environment. Occasional after-hours work is required for activations, upgrades, and major incidents.

Travel

Travel is not a requirement for this role.

Salary Range

$95,000-$110,000

Why Altera?

At Altera Digital Health, you will have the opportunity to profoundly impact the lives of patients by empowering healthcare providers to deliver superior care. You will join a passionate and gifted team committed to innovation and excellence. We offer a competitive compensation and benefits package and the opportunity to work in a fast-paced and dynamic environment.

Key skills/competency

  • Site Reliability Engineering
  • Cloud Computing (Azure)
  • Monitoring and Observability
  • Infrastructure as Code
  • Scripting (PowerShell, Python)
  • Database Administration (SQL Server)
  • Incident Management
  • Networking Fundamentals
  • Windows Server Administration
  • ITSM/ITIL

Skills & topics

  • Site Reliability Engineer
  • SRE
  • Cloud Engineer
  • DevOps Engineer
  • Systems Administrator
  • Azure
  • SQL Server
  • Monitoring
  • Automation
  • IaC
  • PowerShell
  • Python
  • ITSM
  • ITIL
  • Healthcare Technology

How to get hired

  • Tailor your resume: Highlight 7+ years of enterprise support, Azure, SQL Server, and APM tool experience.
  • Showcase automation skills: Emphasize PowerShell, Python, IaC (Terraform), and CI/CD pipeline experience.
  • Address healthcare experience: If applicable, detail any patient care or healthcare technology background.
  • Prepare for technical questions: Be ready to discuss monitoring strategies, incident response, and cloud architecture.
  • Demonstrate ITIL knowledge: Familiarity with ServiceNow and ITIL principles is a plus.

Technical preparation

Master Azure services and networking concepts.,Practice scripting with PowerShell and Python.,Build CI/CD pipelines with Azure DevOps.,Study SQL Server performance tuning techniques.

Behavioral questions

Describe a major incident you resolved.,How do you define and measure SLOs?,How do you collaborate with engineering teams?,How do you handle on-call responsibilities?

Frequently asked questions

What are the primary responsibilities of a Site Reliability Engineer at Altera Digital Health APAC?
As a Site Reliability Engineer at Altera Digital Health APAC, your primary responsibilities will involve ensuring the reliability, scalability, and performance of hosted healthcare platforms. This includes leading incident response, automating operational tasks, defining SLIs/SLOs, and improving monitoring and alerting strategies within cloud and hybrid environments.
What specific technical skills are most important for this Site Reliability Engineer role?
Key technical skills for this Site Reliability Engineer role include extensive experience with APM tools (e.g., Datadog, Dynatrace), deep knowledge of Windows Server administration, strong SQL Server skills, experience with Azure cloud environments, and proficiency in scripting languages like PowerShell or Python. Familiarity with Infrastructure as Code and CI/CD pipelines is also highly valued.
Is this Site Reliability Engineer position remote, and if so, where can candidates be located?
Yes, this Site Reliability Engineer position is fully remote and is open to candidates located within the United States. You will be expected to participate in an on-call rotation to support our 24x7 healthcare environment.
What is the expected salary range for the Site Reliability Engineer position at Altera Digital Health APAC?
The salary range for this Site Reliability Engineer position at Altera Digital Health APAC is between $95,000 and $110,000 annually. The final offer will depend on factors such as experience, skills, and internal equity.
What kind of professional experience is required for the Site Reliability Engineer role?
The role requires a minimum of 7 years of experience supporting enterprise applications, infrastructure, or cloud environments. Strong experience in monitoring and observability, Microsoft stack administration, SQL Server, and Azure cloud environments is essential.
Does Altera Digital Health APAC require a specific degree for the Site Reliability Engineer position?
A Bachelor's Degree in Computer Science, Information Technology, or Engineering is preferred for the Site Reliability Engineer role. However, equivalent professional experience will also be considered, making the role accessible to those with extensive practical backgrounds.
What does 'participate in an on-call rotation' mean for this remote Site Reliability Engineer role?
Participating in an on-call rotation means you will be part of a team that provides support outside of standard business hours to ensure the 24x7 availability of our healthcare systems. This is crucial for maintaining the reliability of our platforms.