
Job Description
Ovii's Interpretation of the Role
The Site Reliability Engineer 2 will design, build, and operate automated release pipelines and Kubernetes‑based infrastructure for Kong's SaaS API platform. The role emphasizes reliability, monitoring, and 24/7 operations across multi‑cloud environments.
Role Snapshot
- Automate CI/CD pipelines for SaaS releases
- Design and manage Kubernetes clusters
- Develop monitoring and alerting systems
- Implement infrastructure as code (Terraform, Chef, etc.)
- Support 24/7 production services
- Deploy across AWS, Azure, GCP and container platforms
Must-Have Requirements
- CI/CD
- Infrastructure as Code (Terraform, Chef, Puppet, Ansible)
- Apache Kafka
- Linux/Unix
- Kubernetes
- Programming languages (Go, C/C++, Python)
- Network services (DNS, TLS/SSL, HTTP)
- Monitoring/alerting systems for API services
- 24/7 service environment experience
- Infrastructure as Code
- 24/7 service operations
- Monitoring/alerting for API services
- BS degree in Computer Science or similar technical field
- Must be located in Singapore
- BS degree in Computer Science or equivalent
Nice-to-Have Signals
- Secure, highly available distributed systems design
- Cloud networking solutions (AWS Transit Gateway, Direct Connect, VPC peering, VPN; Azure VNet; GCP Network Connectivity Center)
- Production software management in AWS
- PostgreSQL multi‑region configuration
- Datadog, ElasticSearch, Prometheus, Grafana
- Redis multi‑region configuration
- Any additional tasks required by manager
- Secure distributed systems
- Cloud networking
- Multi‑region databases
- Observability tooling
Work Setup
- Location: Singapore, Singapore
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Certifications
- Relocation
- Travel
- Security clearance
- Coding test
What You'll Likely Work On
- Create and improve automated release pipelines for SaaS products
- Build, scale, and troubleshoot Kubernetes container clusters
- Develop and maintain monitoring, alerting, and observability tooling for API services
- Implement infrastructure as code using Terraform, Chef, Puppet, or Ansible
- Support 24/7/365 production environments across multi‑cloud infrastructure
- Collaborate with engineering teams to optimize deployment workflows
Good Fit If You Have
- Strong Linux/Unix administration experience
- Hands‑on experience with CI/CD and IaC tools
- Familiarity with cloud networking and Kubernetes troubleshooting
- Experience in high‑availability, distributed systems
- Comfort working in a 24/7 on‑call environment
Skills
- CI/CD
- Infrastructure as Code (Terraform, Chef, Puppet, Ansible)
- Kubernetes
- Linux/Unix
- Programming (Go, C/C++, Python)
- Apache Kafka
- Network services (DNS, TLS/SSL, HTTP)
- Monitoring tools (Datadog, Prometheus, Grafana, ElasticSearch)
- Cloud platforms (AWS, Azure, GCP)
- Distributed systems design