
Senior Engineering Manager, Services Enablement
Hinge Health
Posted 2026-05-08
USD 251,200 - USD 376,800 per year
Tech & Engg
Job Description
Ovii's Interpretation of the Role
Lead the platform reliability and developer enablement team at Hinge Health, ensuring cloud infrastructure stability while driving AI‑native engineering workflows. This senior engineering manager role blends hands‑on SRE expertise with people leadership to deliver high‑availability, cost‑effective services for a fast‑growing healthcare SaaS platform.
Role Snapshot
- Lead platform reliability & scalability
- Manage 8‑10 SRE engineers
- Drive AI‑native workflow initiatives
- Own vendor relationships & cost optimization
- Champion operational excellence & on‑call model
- Build an inclusive high‑performing team
Must-Have Requirements
- AWS
- Kubernetes/EKS
- Terraform
- Platform reliability & SLOs
- Incident management
- Cost optimization
- Cross‑functional leadership
- SaaS platforms
- large‑scale distributed systems
- platform reliability
- SRE
- Bachelor’s Degree in Computer Science, Engineering, or related technical field
Nice-to-Have Signals
- AI‑native workflow tools (Cursor, Claude Code, Copilot)
- Developer productivity metrics (DXI, DORA)
- CI/CD platforms (GitHub Actions, Helm, Nx)
- Observability (Datadog)
- Secrets management (Vault/Infisical)
- Healthcare compliance (HIPAA, HITRUST, SOC2, CCPA)
- NestJS/TypeScript
- AI‑native workflow adoption
- developer productivity improvements
- healthcare compliance
- Master’s Degree in Computer Science, Engineering, or related technical field
Work Setup
- Location: San Francisco, USA
- Work mode: HYBRID
- Remote scope: UNSPECIFIED
- Employment type: Full-Time
Eligibility Gates
- Visa sponsorship: unknown
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Travel
What You'll Likely Work On
- Audit current infrastructure, on‑call practices, and CI/CD pipelines to identify high‑leverage improvements
- Guarantee 99.9%+ uptime across multiple EKS clusters while optimizing costs for AI workloads
- Define safety rails and test harnesses that let autonomous AI agents manage production infrastructure
- Accelerate migration of NestJS services into a monorepo using Nx, GitHub Actions, and Okteto
- Own vendor strategy with AWS, Datadog, Cloudflare, Temporal, and Infisical to drive cost savings
- Run a follow‑the‑sun support model, lead incident reviews, and continuously improve SLOs, runbooks, and observability
Good Fit If You Have
- Thrives at the intersection of platform reliability and AI‑driven operations
- Communicates effectively across engineering, security, product, and executive stakeholders
- Adopts a learn‑it‑all mindset and leads blameless retrospectives
Skills
- AWS cloud infrastructure
- Kubernetes/EKS container orchestration
- Terraform IaC
- CI/CD (GitHub Actions, Helm, Nx)
- Observability (Datadog)
- Secrets management (Vault/Infisical)
- Platform reliability & SLOs
- Cost optimization
- Developer productivity metrics (DXI, DORA)
- AI‑assisted coding tools (Cursor, Claude Code, Copilot)