Senior Software Engineer, SRE

ApplyApply
Posted about 3 hours ago
Share
Renton, Washington
$90 Hourly

Senior Site Reliability Engineer / Platform Engineer

Role Overview

We are seeking a hands-on engineer to build and operate reliable cloud infrastructure for modern software and AI applications. You will improve how teams deploy, secure, observe, and scale production systems while helping establish strong engineering and operational practices.

Responsibilities

  • Design and maintain cloud infrastructure using Terraform and Infrastructure as Code best practices.

  • Build and improve GitOps and CI/CD workflows using tools such as Argo CD or Flux CD.

  • Operate Kubernetes clusters, including upgrades, node pools, autoscaling, and capacity planning.

  • Design and troubleshoot cloud and Kubernetes networking, including private connectivity, ingress, and network segmentation.

  • Implement workload identity and least-privilege access to reduce reliance on static credentials.

  • Define service level objectives, improve observability, and lead incident reviews and root cause analysis.

  • Partner with engineering, product, and security teams to make sound platform decisions.

  • Mentor engineers through design reviews, code reviews, and operational guidance.

Qualifications

  • 5+ years of experience in site reliability, platform, infrastructure, or DevOps engineering.

  • Strong experience with a major cloud platform and Terraform.

  • 3+ years of experience administering production Kubernetes environments.

  • Experience implementing GitOps or automated deployment workflows.

  • Deep understanding of cloud networking and experience resolving production issues.

  • Experience improving the reliability, security, and performance of distributed systems.

  • Ability to explain technical tradeoffs clearly to both technical and nontechnical stakeholders.

Experience supporting AI platforms

Apply