Summary
What you’ll impact
Our company seeks an experienced Cloud DevOps Engineer to design, build, and maintain scalable, secure AWS infrastructure and support Azure workloads. The role involves extensive work with Kubernetes, CI/CD pipelines, GitHub Enterprise, and Terraform, collaborating with engineering and security teams to drive automation and reliability. The position is on‑site in Torrance, CA.
Responsibilities
What you'll do
- AWS: Design, implement, and maintain cloud infrastructure primarily on AWS (EC2, S3, VPC, IAM, RDS, Lambda, EKS, etc.), with secondary support for Azure services
- Kubernetes: Build, manage, and optimize Kubernetes clusters (EKS and/or AKS) for containerized workloads, including scaling, networking, and security configurations
- CI/CD: Design and maintain CI/CD pipelines to automate build, test, and deployment processes across multiple environments
- GitHub Enterprise: Manage source control workflows, branch policies, and repository administration
- Terraform: Write, maintain, and review Terraform modules to provision and manage infrastructure as code across cloud providers
- Implement monitoring, logging, and alerting solutions to ensure system reliability and rapid incident response (e.g., CloudWatch, Prometheus, Grafana, Azure Monitor)
- Collaborate with development teams to improve deployment velocity, reduce lead time, and support GitOps practices
- Enforce and improve cloud security best practices, including IAM policies, secrets management, and compliance standards
- Troubleshoot production issues across cloud infrastructure, container orchestration, and CI/CD pipelines
- Participate in on-call rotation and incident response as needed
- Document infrastructure architecture, runbooks, and operational procedures
- Continuously evaluate and recommend new tools, technologies, and processes to improve system reliability and developer experience
Requirements
What you’ll bring
- 5+ years of experience in a DevOps, Site Reliability Engineering (SRE), or Cloud Infrastructure Engineering role
- AWS: Strong hands-on experience with core services (EC2, VPC, IAM, S3, RDS, EKS, Lambda, CloudFormation, etc.)
- Azure: Working knowledge (AKS, Azure DevOps, VNets, Azure AD, etc.)
- Kubernetes: Solid experience administering and troubleshooting clusters in production environments
- CI/CD: Proven experience building and maintaining pipelines (e.g., GitHub Actions, Jenkins, GitLab CI, or similar)
- GitHub Enterprise: Hands-on experience including repository management, branch protection, and access controls
- Terraform: Strong proficiency for infrastructure-as-code, including module design and state management
- Proficiency in scripting languages such as Bash, Python, or Go
- Solid understanding of networking concepts (DNS, load balancing, VPNs, firewalls)
- Experience with containerization technologies (Docker) and container registries
- Familiarity with monitoring and observability tools (Prometheus, Grafana, CloudWatch, Datadog, or similar)
- Strong understanding of security best practices in cloud environments (IAM, secrets management, least privilege access)
- Applicants must be legally authorized to work in the United States as a U.S. citizen or lawful permanent resident (green card holder).
- This position is fully on-site at our Torrance, CA location. Relocation assistance is not available.