
Sr. Site Reliability Engineer
AuthZed · United States
- Hybrid
- Full-time
- $150,000 / year
- United States
Job highlights
- Ensure system reliability, availability, and performance.
- Design, implement, and maintain scalable infrastructure.
- Automate deployment and configuration processes.
- Collaborate with engineering on resilient systems.
- Troubleshoot and resolve complex production issues.
About the role
About AuthZed:
AuthZed is the creator and maintainer of SpiceDB, providing essential authorization infrastructure for companies worldwide. As a Series A company, we are focused on fixing broken access control with products that simplify permission management while delivering enterprise-scale performance. AuthZed is a fully remote company with employees across the US, Canada, and Europe. We foster a software-driven culture based on integrity, collaboration, and open-mindedness, where every voice is valued.
About the Role:
As a Site Reliability Engineer, you will be instrumental in ensuring the reliability, availability, and performance of our systems. You will design, implement, and maintain scalable infrastructure to support our growing customer base. This is an exciting opportunity to work in a fast-paced, remote environment and contribute to a company bringing a Google-inspired authorization system to a global market.
What you’ll own:
- Design, implement, and maintain highly available and scalable infrastructure solutions.
- Monitor and analyze system performance, resolving bottlenecks to ensure optimal performance and reliability.
- Automate infrastructure deployment and configuration management processes.
- Continuously improve system reliability, security, and efficiency.
- Troubleshoot and resolve complex infrastructure and application issues.
- Collaborate with software engineering teams on resilient, scalable, and secure system design.
- Participate in on-call rotation and respond to production incidents.
- Document system configurations and operational guidelines.
What you bring:
- Proven experience as a Site Reliability Engineer or similar role.
- Strong understanding of networking, operating systems, and cloud infrastructure.
- Experience with Site Reliability Engineering, System Design, and Distributed Computing.
- Proficiency in programming languages like NodeJS, Java, Python, Ruby, and Go.
- Experience with containerization technologies (Docker, Kubernetes).
- Knowledge of infrastructure-as-code tools (Terraform, Pulumi).
- Familiarity with monitoring and logging tools (Prometheus, Grafana, ELK stack).
- Experience with relational databases, especially distributed SQL databases (bonus for Spanner or CockroachDB).
- Experience with Git and GitHub.
- Experience with CI/CD systems.
- Strong problem-solving and troubleshooting skills.
- Excellent communication and collaboration abilities.
Extra shine:
- Experience with Authorization systems.
Life at AuthZed:
- Work with cutting-edge technology in a rapidly growing sector.
- A supportive environment where your ideas drive impact.
- Competitive salary and stock options.
- Comprehensive benefits including healthcare (US-based) and insurance.
- Fully remote work with a flexible schedule.
- Twice-yearly team offsites for bonding and collaboration.
Key skills/competency:
- Site Reliability Engineering
- System Design
- Distributed Computing
- Cloud Infrastructure
- Kubernetes
- Docker
- Terraform
- Prometheus
- Grafana
- Go
Skills & topics
- Site Reliability Engineer
- SRE
- Cloud Engineer
- DevOps
- Infrastructure Engineer
- Kubernetes
- Docker
- Terraform
- Prometheus
- Grafana
- Go
- Python
- System Design
- Distributed Systems
- Remote
How to get hired
- Tailor your resume: Highlight your SRE experience, cloud infrastructure, and distributed systems knowledge, aligning keywords with the job description.
- Showcase technical skills: Emphasize your experience with containerization (Docker, Kubernetes), IaC (Terraform), and monitoring tools (Prometheus, Grafana).
- Demonstrate problem-solving: Prepare to discuss how you’ve troubleshooted complex infrastructure issues and collaborated with engineering teams.
- Understand AuthZed's mission: Research SpiceDB and AuthZed's role in authorization to articulate your interest and fit with their culture and product.
Technical preparation
Behavioral questions
Frequently asked questions
- What are the key responsibilities for a Senior Site Reliability Engineer at AuthZed?
- As a Senior Site Reliability Engineer at AuthZed, you'll be responsible for designing, implementing, and maintaining highly available and scalable infrastructure, monitoring system performance, automating deployments, and collaborating with engineering teams to ensure system resilience and security. You'll also participate in on-call rotations to address production incidents.
- What technical skills are most important for this Senior Site Reliability Engineer role at AuthZed?
- AuthZed highly values experience in Site Reliability Engineering, System Design, and Distributed Computing. Proficiency with containerization (Docker, Kubernetes), infrastructure-as-code (Terraform, Pulumi), programming languages (Go, Python, Java, NodeJS, Ruby), and monitoring tools (Prometheus, Grafana, ELK stack) are critical. Experience with distributed SQL databases is a plus.
- Is this Senior Site Reliability Engineer position at AuthZed fully remote?
- Yes, AuthZed is a fully remote company, and this Senior Site Reliability Engineer position allows for flexible work arrangements across different time zones.
- What kind of company culture can I expect at AuthZed?
- AuthZed fosters a software-driven culture that values Agency, Collaboration, and Open-mindedness. They are a close-knit group that respects diverse perspectives and empowers employees to drive change. Even non-technical teams have a strong understanding and appreciation for the technology.
- What opportunities for growth are available for a Senior Site Reliability Engineer at AuthZed?
- As an early-stage startup, AuthZed offers the opportunity to work with cutting-edge technology in a rapidly growing sector. Your ideas will lead to real impact, and you'll be part of a team shaping the future of authorization infrastructure.
Similar roles
Open positions we recommend based on this role.