
Software Engineer II - Cloud Infrastructure Engineer
Abnormal AI · United States
- Hybrid
- Full-time
- $180,000 / year
- United States
Job highlights
- Engineer cloud infrastructure for microservices.
- Manage cell lifecycle and deployment tooling.
- Design scalable, cell-native architecture.
- Maintain secure, cost-efficient multi-cloud infra.
- Resolve complex cross-layer issues.
About the role
About The Role
Abnormal AI is an AI-native behavioral security platform that protects enterprises from advanced threats by analyzing and understanding communication patterns and access behavior at scale. We now protect more than 25% of the Fortune 500, and as we expand into new product lines and geographies, a scalable, reliable infrastructure foundation is critical to our next phase of growth.
The Platform & Infrastructure team is seeking a Cloud Infrastructure Engineer for our Cellular Infrastructure team. This team owns the full lifecycle of Abnormal’s cell-based deployment architecture—bootstrapping new cells, deploying our entire application and infrastructure stack onto them, and keeping every cell healthy, isolated, cost-efficient, and compliant. Engineers on this team wear multiple hats: infra engineering, application-layer debugging, and close collaboration with product and application teams to minimize overhead so those teams can stay focused on building.
What You Will Do
- Bootstrap new cells end-to-end: full infrastructure setup (compute, networking, IAM, etc.) and complete application stack deployment.
- Maintain and evolve cell lifecycle tooling to make provisioning repeatable, auditable, and operator-friendly—reducing manual steps and time-to-production.
- Partner with application and product teams to design and implement scalable, cell-native architecture approaches.
- Design, build, test, scale, monitor, and maintain secure, cost-efficient infrastructure in a multi-cloud environment (AWS and Azure).
- Triage and resolve complex cross-layer issues quickly, then drive root cause fixes that prevent recurrence.
- Drive down technical debt and toil through automation and systemic improvements to the cell deployment lifecycle.
- Participate in on-call rotation with a learning-oriented mindset, identifying systemic gaps and driving long-term reliability improvements.
- Keep cross-team communication low-friction and high-signal: proactive and well-contextualized.
- Contribute as a core member of an agile team through sprint planning, standups, and execution with a strong sense of ownership and teamwork.
Must Haves
- Bachelor’s degree in Computer Science or a related technical field.
- 4+ years of experience engineering cloud infrastructure for production microservice systems, with attention to performance, reliability, security, and cost.
- 2+ years of Python experience, including application-layer code (not just scripts).
- 1+ year of experience with Kubernetes and Helm.
- 1+ year of AWS experience ( VPC, IAM, S3, Route 53, CloudFront, EKS, ECS, CloudWatch).
- 1+ year of Terraform and HCL experience.
- Comfort operating across infra and application engineering without hard boundaries.
- Experience with on-call rotations, incident response, and operating production-grade systems.
- Practical experience using Generative AI tools in day-to-day engineering workflows.
- Strong communication skills and the ability to thrive in a fast-paced, remote-first environment—balancing autonomy with collaboration, demonstrating a bias toward action, and maintaining a positive, constructive mindset.
Nice to Haves
- Experience with Bash, Golang, Terragrunt and data infrastructure (Spark, Databricks).
- Hands-on experience with cell-based, multi-tenant, or multi-region infrastructure architectures.
- Familiarity with Generative AI developer tools such as Claude Code, and experience driving AI-first engineering workflows.
- Prior experience building large-scale IaC abstractions or internal developer platforms.
- AWS certifications.
Compensation & Benefits
Actual compensation will be determined based on several non-discriminatory factors including skills, experience, qualifications, and geographic location.
In addition to base salary, this role may be eligible for bonus or incentive compensation, equity, and a comprehensive benefits package.
Base salary range: $149,200—$214,500 USD
Equal Opportunity Employer
Abnormal AI is an equal opportunity employer. Qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability, protected veteran status or other characteristics protected by law. For our EEO policy statement please click here. If you would like more information on your EEO rights under the law, please click here.
Key skills/competency
- Cloud Infrastructure Engineer
- AWS
- Azure
- Kubernetes
- Terraform
- Python
- IaC
- Microservices
- Security
- Reliability
Skills & topics
- Cloud Infrastructure Engineer
- AWS
- Azure
- Kubernetes
- Terraform
- Python
- IaC
- Microservices
- Security
- Reliability
- DevOps
- SRE
- Site Reliability Engineering
- Infrastructure as Code
- System Design
- Remote Work
- AI
- Machine Learning Infrastructure
How to get hired
- Tailor your resume: Highlight your 4+ years of cloud infrastructure experience, Python, Kubernetes, AWS, and Terraform skills, aligning them with the 'Must Haves' section.
- Showcase AI proficiency: Emphasize any practical experience with Generative AI tools in your application and during interviews for this Cloud Infrastructure Engineer role.
- Prepare for technical questions: Be ready to discuss your experience with microservices, IaC, on-call rotations, and incident response for production systems.
- Demonstrate collaboration: Highlight your communication skills and ability to thrive in a fast-paced, remote-first environment, emphasizing teamwork and ownership.
- Research Abnormal AI: Understand their AI-native behavioral security platform, their mission, and their growth, showing genuine interest in the company and the Cloud Infrastructure Engineer position.
Technical preparation
Behavioral questions
Frequently asked questions
- What is the base salary range for a Cloud Infrastructure Engineer at Abnormal AI?
- The base salary range for the Cloud Infrastructure Engineer position at Abnormal AI is $149,200 to $214,500 USD annually. Actual compensation depends on factors like skills, experience, qualifications, and location.
- What are the primary responsibilities of a Cloud Infrastructure Engineer at Abnormal AI?
- As a Cloud Infrastructure Engineer at Abnormal AI, you will be responsible for bootstrapping new cells, maintaining cell lifecycle tooling, partnering with product teams on architecture, designing and maintaining secure, cost-efficient infrastructure in AWS and Azure, triaging and resolving cross-layer issues, and reducing technical debt through automation.
- What technical skills are essential for the Cloud Infrastructure Engineer role at Abnormal AI?
- Essential technical skills include a Bachelor's degree in Computer Science or related field, 4+ years of cloud infrastructure experience for microservices, 2+ years of Python, 1+ year of Kubernetes/Helm, 1+ year of AWS, and 1+ year of Terraform/HCL. Experience with on-call rotations and Generative AI tools is also required.
- Does Abnormal AI offer remote work opportunities for the Cloud Infrastructure Engineer position?
- Yes, Abnormal AI emphasizes a remote-first environment for this Cloud Infrastructure Engineer role, requiring strong communication skills and the ability to balance autonomy with collaboration.
- What are the 'Nice to Haves' for an applicant applying for the Cloud Infrastructure Engineer role?
- Nice-to-have skills include experience with Bash, Golang, Terragrunt, data infrastructure (Spark, Databricks), cell-based/multi-tenant/multi-region architectures, Generative AI developer tools, building large-scale IaC abstractions, and AWS certifications.
- How does Abnormal AI approach career growth for its Cloud Infrastructure Engineers?
- While not explicitly detailed, the role description highlights a learning-oriented mindset for on-call rotations and driving long-term reliability improvements, suggesting opportunities for skill development and advancement within the Platform & Infrastructure team.
- What is the company culture like at Abnormal AI for a Cloud Infrastructure Engineer?
- Abnormal AI fosters a culture that is AI-native, focused on behavioral security, and values a scalable, reliable infrastructure foundation. The Platform & Infrastructure team operates in an agile manner with a strong sense of ownership and teamwork, emphasizing low-friction, high-signal cross-team communication in a remote-first setting.