
Job Description
Ovii's Interpretation of the Role
Kong is hiring a Site Reliability Engineer in Milan to design, build, and operate the cloud infrastructure that powers its API and AI platform. The role blends infrastructure‑as‑code, automation, and reliability engineering to keep services running at 99.99% uptime while supporting fast product delivery.
Role Snapshot
- Hybrid SRE role based in Milan
- Build and manage cloud infrastructure as code
- Drive reliability, performance, and uptime
- Automate operations and reduce toil
- Collaborate with engineering teams
- Participate in on‑call rotation
Must-Have Requirements
- Cloud platform operations (AWS, GCP, Azure)
- Programming/scripting (Golang, Python, Bash)
- Containerization and orchestration (Docker, Kubernetes)
- Operating production workloads on major cloud providers
- Containerization and orchestration
- Programming/scripting language proficiency
Nice-to-Have Signals
- Terraform (IaC)
- Ansible
- CI/CD tools (GitLab CI, Jenkins)
- Observability tools (Prometheus, Grafana, ELK)
- Infrastructure as Code (Terraform)
- CI/CD pipeline familiarity
- Observability stack knowledge
Work Setup
- Location: Milan, Italy
- Work mode: HYBRID
- Employment type: Full-Time
Not Specified in JD
- Salary range
- Visa sponsorship
- Relocation
- Travel
- Security clearance
- Coding test
- Portfolio
- GitHub
What You'll Likely Work On
- Develop and maintain infrastructure as code using Terraform and Ansible
- Implement monitoring, logging, and alerting to achieve 99.99% uptime
- Handle production incidents and lead blameless post‑mortems
- Automate operational tasks and enable self‑service for engineering teams
- Partner with developers to embed reliability and scalability practices
- Contribute to capacity planning, disaster‑recovery drills and security hardening
- Take part in a sustainable on‑call rotation
Good Fit If You Have
- Experience with Terraform or Ansible (preferred)
- Familiarity with CI/CD tools such as GitLab CI or Jenkins (optional)
- Knowledge of observability stacks like Prometheus, Grafana, or ELK (optional)
Skills
- Cloud platforms (AWS, GCP, Azure)
- Infrastructure as Code (Terraform, Ansible)
- Container orchestration (Docker, Kubernetes)
- Programming (Go, Python, Bash)
- CI/CD pipelines (GitLab CI, Jenkins)
- Observability (Prometheus, Grafana, ELK)