Ovii Job Board

Senior SRE, Managed Gateways

Kong

Olympia, United States • Remote - Olympia, United States • Full-Time

Posted 2026-07-23 USD 113,000 - USD 162,000 per year Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

Kong seeks a Senior Site Reliability Engineer to design, operate, and scale the Managed Gateways SaaS platform. You will own reliability for a multi‑cloud service and serve as the technical lead for enterprise implementations worldwide.

Role Snapshot

  • Senior Site Reliability Engineer
  • Multi‑cloud (AWS, GCP, Azure) focus
  • Kubernetes‑native architecture
  • Automation & tooling (Go, Terraform, Ansible)
  • Enterprise customer implementation

Must-Have Requirements

  • Site Reliability Engineering
  • Kubernetes
  • Go (or similar language)
  • Terraform
  • Ansible
  • CI/CD pipeline development
  • Monitoring/Logging/Alerting (Prometheus, Grafana, ELK, Datadog)
  • Site reliability engineering
  • Kubernetes and cloud‑native system design
  • Automation and infrastructure as code
  • Monitoring, alerting, and incident management

Nice-to-Have Signals

  • Multi‑cloud experience (AWS, GCP, Azure)
  • Service Mesh (Istio, Linkerd)
  • Database administration (PostgreSQL, Cassandra)
  • Open‑source SRE tool contributions
  • Relevant cloud certifications (AWS DevOps Engineer, CKA)
  • Multi‑cloud deployments
  • Service mesh technologies
  • Database operations for high‑throughput systems
  • Open‑source contributions

Work Setup

  • Location: Washington, United States
  • Work mode: REMOTE
  • Remote scope: UNSPECIFIED
  • Remote countries: United States, Washington, United States
  • Employment type: Full-Time

Eligibility Gates

  • Visa sponsorship: unknown

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility details
  • Education requirement
  • Certification requirement
  • Relocation
  • Travel
  • Security clearance
  • Coding test

What You'll Likely Work On

  • Architect and operate resilient, fault‑tolerant cloud‑native systems for Managed Gateways
  • Lead incident response, blameless post‑mortems, and continuous service improvement
  • Define, track, and report SLOs/SLIs to ensure platform reliability
  • Build automation, self‑service tooling, and CI/CD pipelines using Go, Terraform, and Ansible
  • Mentor and guide a team of SREs while influencing product road‑maps
  • Partner with enterprise customers to design, deploy, and hand‑off production‑ready gateway solutions
  • Create repeatable implementation playbooks and feed customer insights back to product

Good Fit If You Have

  • Thrives in fast‑moving, distributed teams across time zones
  • Shows strong ownership and urgency in critical incidents
  • Enjoys sharing knowledge and elevating team capabilities

Skills

  • Site Reliability Engineering
  • Kubernetes
  • Go (or similar language)
  • Infrastructure as Code (Terraform, Ansible)
  • CI/CD pipelines
  • Monitoring & observability (Prometheus, Grafana, ELK, Datadog)
  • Multi‑cloud architecture
  • Service Mesh (Istio, Linkerd) – preferred
  • Database ops (PostgreSQL, Cassandra) – preferred
  • Open‑source SRE contributions – preferred

Remote Eligibility

  • United States
  • Washington, United States