Featured Job

Staff Site Reliability Engineer- Developer Platform

Palo Alto, CA Full-time On-site $186k — $255.8k per year 08/24/2026 Job ID: 000029
Apply Now
Site Reliability Engineering Platform Engineering DevOps Infrastructure as Code Terraform

Summary

What you’ll impact

The role is for an experienced Site Reliability Engineer who will design, build, and operate infrastructure supporting build pipelines for firmware delivery. The engineer will work on cloud‑native systems, developer platforms, and DevOps culture across a large developer organization.

Responsibilities

What you'll do

  • Contribute to the design and implementation of infrastructure components, from a developer platform mindset, including centralized developer portals, tooling to reduce developer cognitive overhead and
  • Help manage our AWS footprint by identifying opportunities for better multi-region availability, disaster recovery, and cost efficiency.
  • Participate in defining Infrastructure as Code (IaC) patterns and help maintain the library of modules used across the engineering organization.
  • Develop and optimize custom Kubernetes operators and controllers to automate stateful services, ensuring high availability and performance.
  • Assist in building robust, automated security controls and policy-as-code to ensure platform compliance.
  • Identify friction points across the organizations that operate on top of our infrastructure and collaboratively build intuitive tools that improve the developer experience.
  • Support mid-level and junior engineers through code reviews, pair programming, and technical leadership.
  • Work closely with engineering teams to gather requirements, identify dependencies, and provide technical input on risk mitigation.

Requirements

What you’ll bring

  • 5+ years of relevant experience in Platform Engineering, DevOps, or SRE roles.
  • Proficiency with Infrastructure as code, preferably Terraform. Experience designing reusable modules and managing state at scale.
  • Deep understanding of Kubernetes internals, networking, storage and service mesh architectures.
  • Proven track record implementing GitOps workflows using ArgoCD or Flux in production environments.
  • Strong understanding of automation and scripting (e.g., Python, Bash, GoLang).
  • Strong experience with one of the major Cloud Service Providers (AWS, Azure, or GCP), including IAM, compute, networking and storage layers.
  • Experience maintaining and optimizing CI pipelines.
  • Ability to clearly communicate technical concepts to peers and stakeholders.
  • Comfortable working in a collaborative environment where architectural decisions are discussed and debated.
  • Willingness to mentor peers and share technical knowledge.

Ready to Move Forward?

Apply now and our recruiting team will reach out with next steps, interview guidance, and client insights tailored to this role.