Ovii Job Board

Software Engineer, DevOps

Ema

San Francisco Bay Area, USA • Onsite - San Francisco Bay Area, USA • Full-Time • 3+ years

Posted 2026-04-30 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

Ema is looking for a DevOps Engineer to design, build, and operate scalable, multi‑tenant SaaS infrastructure. The role blends cloud platform expertise, automation, CI/CD pipeline development, and observability to keep the Agentic AI platform reliable and performant.

Role Snapshot

  • Design and build scalable cloud infrastructure
  • Automate provisioning and configuration
  • Develop and maintain CI/CD pipelines
  • Implement observability and alerting
  • Collaborate with product and engineering teams
  • Support multi‑tenant SaaS services

Must-Have Requirements

  • Kubernetes
  • Terraform
  • Helm
  • Istio
  • AWS/Azure/GCP
  • Ansible
  • GitHub Actions
  • Prometheus
  • Grafana
  • Sentry
  • PagerDuty
  • Selenium
  • Infrastructure engineering
  • cloud platforms
  • CI/CD pipeline development
  • Bachelor's or Master's degree in Computer Science or related field

Nice-to-Have Signals

  • ML/OPs
  • PostgreSQL query optimization
  • Event‑driven data pipelines
  • Air‑gapped cloud environments
  • Azure AKS
  • ML/Ops
  • PostgreSQL performance tuning
  • event‑driven pipelines

Work Setup

  • Location: San Francisco Bay Area, United States
  • Work mode: ONSITE
  • Employment type: Full-Time

Not Specified in JD

  • Visa sponsorship
  • Remote eligibility
  • Travel
  • Security clearance
  • Coding test
  • Portfolio
  • GitHub
  • Writing sample
  • Cover letter

What You'll Likely Work On

  • Partner with product teams to architect and build foundational infrastructure for the AI platform.
  • Create highly available, multi‑tenant SaaS solutions on AWS, Azure, or GCP using Kubernetes, Helm, Terraform, and Istio.
  • Automate infrastructure tasks from provisioning to configuration management with Terraform, Ansible, and Kubernetes.
  • Enhance CI/CD pipelines using GitHub Actions, Cloud Build, and related tooling to improve developer experience.
  • Build observability stacks with Prometheus, Grafana, Sentry, and PagerDuty for real‑time health monitoring.
  • Deploy end‑to‑end testing frameworks such as Selenium to ensure software quality.
  • Monitor system performance, analyze metrics, and optimize database queries for peak efficiency.

Good Fit If You Have

  • Experience with machine‑learning operations or ML‑focused pipelines.
  • Strong knowledge of PostgreSQL query tuning and performance improvement.
  • Background in event‑driven architectures or streaming data pipelines.
  • Familiarity with air‑gapped or private cloud environments.
  • Hands‑on experience administering Azure Kubernetes Service (AKS).

Skills

  • Kubernetes
  • Terraform
  • Helm
  • Istio
  • Public cloud platforms (AWS, Azure, GCP)
  • Ansible
  • GitHub Actions / Cloud Build
  • Prometheus & Grafana
  • Sentry & PagerDuty
  • Selenium testing
  • PostgreSQL query optimization (preferred)
  • ML/OPs (preferred)