Job Description
Ovii's Interpretation of the Role
The DevOps Engineer will own the reliability, deployment speed, and automation of Sarvam's Studio media platform. You will manage production Kubernetes clusters, CI/CD pipelines, observability, and cost‑optimization across a multi‑cloud environment.
Role Snapshot
- Production Kubernetes cluster ownership
- CI/CD pipeline design & automation
- Multi‑cloud infrastructure management
- Observability & incident response
- Cost‑optimization and resource sizing
Must-Have Requirements
- Kubernetes (production)
- Helm
- Azure/GCP/AWS
- CI/CD pipelines
- Docker
- Prometheus/Grafana (observability)
- Linux systems
- Python or Bash scripting
- DevOps
- Site Reliability Engineering
- Infrastructure Engineering
Nice-to-Have Signals
- Multi‑cloud experience
- GPU workloads on Kubernetes
- ML model serving (Triton, TorchServe, vLLM)
- Service mesh (Linkerd, Istio, Consul)
- GitOps (ArgoCD, Flux)
- Event‑driven autoscaling (KEDA)
- Media/audio/video processing pipelines
- Infrastructure‑as‑code (Terraform, Pulumi)
- Multi‑tenant Kubernetes clusters
- Cloud certifications (Azure, GCP, AWS)
- Redis or similar in‑memory data stores
- multi‑cloud environments
- GPU workloads
- ML model serving
- Azure
- GCP
- AWS
Work Setup
- Location: Bengaluru, India
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Education requirement
- Certifications
- Relocation
- Notice period
- Travel
- Security clearance
- Coding test
- Portfolio
- GitHub
- Writing sample
- Cover letter
What You'll Likely Work On
- Operate and scale production Kubernetes clusters, including HPAs, CronJobs, and node pool scheduling
- Design, build, and maintain CI/CD pipelines with automated testing, container image builds, and staged rollouts
- Create and manage Helm‑based deployment charts and environment‑specific value overlays
- Implement observability across services using metrics, dashboards, tracing, and alerting
- Manage cloud resources such as blob storage, CDN, ingress, IAM policies, and secrets vaults
- Coordinate multi‑cloud deployments, artifact management, and credential handling
- Develop tooling for developer productivity, local environments, and self‑service infrastructure
- Lead incident response processes, runbooks, and post‑mortem improvements
Good Fit If You Have
- Enjoys fast‑paced, small‑team environments
- Strong analytical and problem‑solving mindset
- Collaborative communication with engineering and product stakeholders
Skills
- Kubernetes (production)
- Helm charts
- Azure / GCP / AWS
- CI/CD pipelines
- Docker containerization
- Prometheus / Grafana observability
- Linux systems
- Python / Bash scripting
- Multi‑cloud operations (preferred)
- GPU workloads on Kubernetes (preferred)
- Infrastructure‑as‑code (Terraform/Pulumi) (preferred)