Featured Job

Senior Reliability Engineer

Seattle Full-time Remote $120k — $190k per year 09/17/2026 Job ID: 000220
Apply Now
AWS Fly.io Terraform CloudFormation deployment strategy

Summary

What you’ll impact

Our company is seeking a Senior Software Engineer to own infrastructure, reliability, and platform engineering, designing and scaling cloud systems that keep the identity verification platform fast, secure, and highly available. The role involves building tooling to improve developer productivity, collaborating with product and engineering leadership, and ensuring robust observability and incident response practices.

Responsibilities

What you'll do

  • Design, build, and maintain scalable, cost-effective cloud infrastructure across AWS and Fly.io.
  • Own deployment strategy and service design across our monolith and microservices, including schema migrations and rollout safety.
  • Manage infrastructure-as-code (CloudFormation, Terraform or equivalent) for reproducible, auditable infrastructure.
  • Identify reliability risks, performance bottlenecks, and security gaps, and address them proactively.
  • Evolve our observability stack: logging, distributed tracing, metrics, and uptime monitoring.
  • Evolve on-call practices and incident response processes that keep us ahead of customer-impacting issues.
  • Champion a culture of reliability: postmortems, runbooks, and continuous improvement after incidents.
  • Manage and improve our CI/CD pipelines (GitHub Actions) and establish deployment best practices.
  • Build internal platform tooling and abstractions that reduce toil and increase engineering velocity.
  • Partner with product engineers to make infrastructure easy to use correctly and hard to use incorrectly.
  • Design and operate data pipelines that support our ML-powered verification systems.
  • Evolve our MLOps infrastructure so models can be trained, evaluated, and deployed safely and repeatedly.
  • Work closely with engineering and product leadership on technical roadmap decisions.
  • Review code, mentor peers, and help raise the bar on security, reliability, and operational discipline.
  • Communicate infrastructure tradeoffs clearly across technical and non-technical stakeholders.

Requirements

What you’ll bring

  • Work Authorization (Required): Applicants must be legally authorized to work in the United States for any employer without current or future need for visa sponsorship. This is a firm requirement.
  • Cloud & Infrastructure: Hands-on experience managing and securing cloud infrastructure. AWS required; Fly.io or similar a plus.
  • Infrastructure-as-Code: Production experience with Terraform or equivalent tools.
  • Databases: Deep experience with PostgreSQL, including schema design, migrations, and query performance tuning.
  • CI/CD: Strong experience designing and managing pipelines, GitHub Actions preferred.
  • Languages: Proficiency in modern, type-safe languages. Go strongly preferred.
  • Observability: Experience building and operating logging, tracing, and metrics systems in production.
  • MLOps & Data Pipelines: Prior experience shipping ML infrastructure into production, not just experimentation.
  • Startup experience: You've worked at an early-stage company and know what it means to move fast without compromising the things that matter.
  • Security: Security-minded by default. You design with the threat model in mind, not as an afterthought.

Ready to Move Forward?

Apply now and our recruiting team will reach out with next steps, interview guidance, and client insights tailored to this role.