
Site Reliability Engineer
SS&C Technologies · Arizona, United States
- Hybrid
- Full-time
- $120,000 / year
- Arizona, United States
Job highlights
- Lead teams building resilient cloud infrastructure.
- Enhance application availability and reliability.
- Automate processes and eliminate toil.
- Integrate DevSecOps and zero-trust principles.
- Drive continuous improvement with data.
About the role
Site Reliability Engineer (SRE)
SS&C Technologies is a global leader in financial services and healthcare technology, known for its expertise, scale, and innovation. We empower thousands of organizations worldwide with cutting-edge solutions.
About the Role
As a Site Reliability Engineer (SRE) at SS&C Technologies, you will be instrumental in building and operating scalable, resilient, and secure infrastructure platforms and services. This role is pivotal in enabling our global business units to innovate confidently, modernize applications, and manage technical debt effectively. You will foster crucial collaboration between product, engineering, and operations teams, embedding reliability, automation, and compliance into our development lifecycle.
This position is part of the Global Technology Infrastructure SRE team.
What You Will Get To Do
- Collaborate with Technology Infrastructure teams to build and operate cloud-native platforms that abstract complexity and accelerate delivery, ensuring reliability from design to operations.
- Partner with business and technical teams to enhance application availability, observability, and reliability during migrations to the Private Cloud.
- Improve platform reliability through automated problem detection, self-healing systems, and robust notification/escalation protocols.
- Utilize SLOs, SLIs, and KPIs to drive prioritization, measure impact, and guide continuous improvement efforts.
- Eliminate toil by implementing intelligent automation and agentic workflows.
- Conduct blameless retrospectives and disseminate learnings across the organization.
- Cultivate a culture of ownership, positive thinking, and continuous learning, grounded in practicality and engineering excellence.
- Integrate DevSecOps, zero-trust principles, and policy-as-code into all pipelines.
- Produce and promote Architecture Decision Records (ADRs) and Cloud Well-Architected Frameworks for technology standardization.
- Maintain 24x5 active coverage with seamless regional handoffs and weekend escalation protocols.
What You Will Bring
- 5+ years of professional experience in a SRE role, with 3+ years in financial services or other regulated industries preferred.
- Minimum Bachelor’s degree in Computer Science, Engineering, or a related field.
- Proven expertise in architecting, designing, and operating private cloud environments (e.g., VMware, OpenStack, OpenShift Virtualization) and Kubernetes clusters at scale.
- Hands-on experience with building, deploying, and operating infrastructure as code platforms, CI/CD pipelines, and observability tools (e.g., Prometheus, Splunk).
- Strong understanding of modern systems reliability standards, including KPIs, SLAs, SLOs, and deriving actionable insights.
- Familiarity with financial services regulatory frameworks and their infrastructure implications.
- Familiarity with structured naming conventions and asset management for global infrastructure.
- Experience with financial-grade network segmentation, micro-segmentation, and zero-trust architecture.
- Certifications such as TOGAF, AWS Certified Solutions Architect, VMware VCP, or Red Hat Certified Architect are a plus.
- Familiarity with ISO 27001, NIST 800-53, and other security frameworks is a plus.
Our Expectations
- Outstanding organization, project management, and attention to detail, with strong decision-making and problem-solving skills.
- A tenacious problem solver and continuous learner, adaptable to new technologies in a fast-paced environment.
- Powerful verbal and written communication skills, capable of articulating complex concepts and collaborating effectively under pressure.
- Ability to quickly establish credibility with diverse technical stakeholders, including executives and various technical teams.
- Discretion in handling confidential information and adherence to compliance and regulatory requirements.
- Commitment to high professional and ethical standards in a diverse workplace.
- Flexibility to work non-traditional hours and occasional travel as needed.
Why You Will Love It Here!
- Flexibility: Hybrid Work Model & a Business Casual Dress Code.
- Your Future: 401k Matching Program, Professional Development Reimbursement.
- Work/Life Balance: Flexible Personal/Vacation Time Off, Sick Leave, Paid Holidays.
- Your Wellbeing: Medical, Dental, Vision, Employee Assistance Program, Parental Leave.
- Wide Ranging Perspectives: Commitment to Diversity and Inclusion.
- Training: Hands-On, Team-Customized, including SS&C University.
- Extra Perks: Discounts on fitness clubs, travel, and more.
SS&C Technologies is an Equal Employment Opportunity employer committed to diversity and inclusion.
Key skills/competency
- Site Reliability Engineering
- Cloud-Native Platforms
- Kubernetes
- Infrastructure as Code
- CI/CD
- Observability
- SLOs/SLIs/KPIs
- DevSecOps
- Zero-Trust Architecture
- Automation
Skills & topics
- Site Reliability Engineer
- SRE
- Cloud Engineer
- DevOps Engineer
- Infrastructure Engineer
- Reliability Engineering
- Automation
- Kubernetes
- Cloud Computing
- Financial Services Technology
How to get hired
- Tailor your resume: Highlight SRE experience, private cloud, Kubernetes, and financial services background.
- Showcase automation skills: Detail experience with IaC, CI/CD, and observability platforms like Prometheus or Splunk.
- Emphasize reliability standards: Include knowledge of SLOs, SLIs, KPIs, and regulated industry compliance.
- Prepare for technical questions: Be ready to discuss architecture, private cloud, and microservices concepts.
- Demonstrate soft skills: Highlight problem-solving, communication, and collaboration in demanding environments.
Technical preparation
Behavioral questions
Frequently asked questions
- What is the work arrangement for the Site Reliability Engineer role at SS&C Technologies?
- This Site Reliability Engineer position is a remote role, with specific state focuses in FL, TX, GA, NC, AZ, and TN. While the role is remote, SS&C Technologies also offers a hybrid work model for many of its positions, suggesting flexibility is a key aspect of their culture.
- What are the key technical skills required for the Site Reliability Engineer position at SS&C Technologies?
- The Site Reliability Engineer role demands expertise in architecting and operating private cloud environments (VMware, OpenStack, OpenShift Virtualization) and Kubernetes. You'll need hands-on experience with infrastructure as code, CI/CD pipelines, and observability tools like Prometheus and Splunk. A strong understanding of modern systems reliability standards, including SLOs, SLIs, and KPIs, is also crucial.
- Does SS&C Technologies prefer candidates with experience in regulated industries for the SRE role?
- Yes, SS&C Technologies prefers candidates with 3+ years of experience in financial services or other regulated industries for the Site Reliability Engineer position. Familiarity with financial services regulatory frameworks and their impact on infrastructure design and operations is highly valued.
- What are the career growth opportunities for a Site Reliability Engineer at SS&C Technologies?
- SS&C Technologies emphasizes professional development through programs like Professional Development Reimbursement and SS&C University. The company culture encourages continuous learning and provides hands-on, team-customized training, offering ample opportunities for growth within the SRE domain and beyond.
- How does SS&C Technologies approach work-life balance for its Site Reliability Engineers?
- SS&C Technologies promotes work-life balance through a flexible personal/vacation time off policy, sick leave, and paid holidays. They also offer benefits like a 401k matching program and parental leave, alongside a hybrid work model and business casual dress code, aiming to support employee wellbeing.
- What is the importance of automation in the Site Reliability Engineer role at SS&C Technologies?
- Automation is a core aspect of the Site Reliability Engineer role at SS&C Technologies. The job description explicitly mentions eliminating toil using intelligent automation and agentic workflows, and integrating DevSecOps and policy-as-code into pipelines, highlighting its central importance to the role's success.