Ovii Job Board

Staff Engineer, API Platform

Sarvam

Bengaluru, India • Onsite - Bengaluru, India • Full-Time • 7+ years

Posted 2026-05-21 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

The Staff Engineer will own Sarvam’s API Platform, driving its architecture, reliability, and performance. You will lead the rewrite from Python/FastAPI to Go, build streaming and commercial layers, and mentor cross‑functional teams.

Role Snapshot

  • Own end‑to‑end API platform
  • Lead Python‑to‑Go rewrite
  • Design high‑scale streaming infrastructure
  • Build metering, billing and rate‑limiting
  • Ensure sub‑second latency and reliability
  • Mentor inference and MLOps teams

Must-Have Requirements

  • Go
  • Python
  • Kubernetes
  • PostgreSQL
  • Redis
  • WebSockets
  • Distributed systems design
  • production backend systems
  • high‑scale distributed systems
  • low‑latency multi‑tenant platforms

Nice-to-Have Signals

  • LLM serving
  • Audio processing
  • Billing / metering infrastructure
  • Early‑stage startup experience
  • FastAPI
  • LLM / ASR / TTS / vision model serving
  • audio processing pipelines
  • billing and metering systems

Work Setup

  • Location: Bengaluru, India
  • Work mode: ONSITE
  • Employment type: Full-Time

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility
  • Education requirement
  • Certifications
  • Relocation
  • Notice period
  • Travel
  • Security clearance
  • Coding test
  • Portfolio
  • GitHub
  • Writing sample
  • Cover letter

What You'll Likely Work On

  • Define and evolve the platform architecture from edge request to response
  • Drive the Python‑to‑Go migration, setting patterns and ensuring reliability
  • Engineer audio and vision pipelines for ASR, TTS, OCR and VLM use‑cases
  • Build streaming infrastructure with WebSockets, back‑pressure handling and low‑latency model invocation
  • Implement commercial features such as metering, billing, prepaid wallets and rate‑limiting
  • Maintain sub‑second latency SLOs for a multi‑tenant platform
  • Provide observability (logging, metrics, tracing) and robust integration testing
  • Partner with Inference and MLOps teams to embed production engineering best practices

Good Fit If You Have

  • Experience serving LLM, ASR, TTS, or vision models in production
  • Background in audio processing (e.g., FFmpeg, VAD)
  • Prior work building metering, billing, or payment‑related infrastructure
  • Time spent at early‑stage or growth‑stage startups

Skills

  • Go
  • Python
  • FastAPI
  • Kubernetes
  • PostgreSQL
  • Redis
  • WebSockets
  • Streaming systems
  • Observability (metrics, tracing)
  • API design
  • Distributed systems