Ovii Job Board

Artificial Intelligence Engineer

Innovaccer

San Francisco, United States • Onsite - San Francisco, United States • Full-Time

Posted 2026-08-10 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

Innovaccer seeks an AI Engineer to design, develop, and ship production‑grade AI solutions for its healthcare platform. You will partner with product, engineering, and data teams to build LLM‑based applications, AI agents, and retrieval‑augmented generation pipelines from research to deployment.

Role Snapshot

  • Build production‑grade AI models and pipelines
  • Develop LLM‑based applications and AI agents
  • Deploy and optimize AI services at scale
  • Collaborate across product, engineering, and data teams
  • Design experiments and evaluate model performance
  • Write production‑ready Python/PyTorch code

Must-Have Requirements

  • Python
  • PyTorch
  • AI model development
  • LLM solutions
  • Production deployment
  • MS in Computer Science, Machine Learning, or related quantitative field
  • PhD in Computer Science, Machine Learning, or related quantitative field

Nice-to-Have Signals

  • Familiarity with HuggingFace, DeepSpeed/FSDP, vLLM, SGLang
  • Multi‑GPU training experience
  • Knowledge of sharding strategies
  • Open‑source ML contributions
  • First‑author publications at top conferences
  • Some exposure to multi‑GPU training
  • Research publications
  • Open‑source contributions
  • BS with substantial research or open‑source work (exceptional candidates)

Work Setup

  • Location: San Francisco, United States
  • Work mode: ONSITE
  • Employment type: Full-Time

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility
  • Certification requirements
  • Shift or travel expectations

What You'll Likely Work On

  • Design, prototype, and ship AI models and pipelines that power healthcare applications
  • Implement LLM‑based solutions, AI agents, and retrieval‑augmented generation workflows
  • Fine‑tune open‑weight models, choose parameter‑efficient or full fine‑tuning, and meet product‑level accuracy targets
  • Build and maintain training and serving infrastructure using HuggingFace, DeepSpeed/FSDP, vLLM/SGLang, or equivalents
  • Collaborate with product, engineering, and data teams to translate research ideas into production features
  • Communicate model behavior and results to clinicians and other non‑technical stakeholders

Good Fit If You Have

  • Strong curiosity for cutting‑edge AI research
  • Ability to explain technical concepts to non‑technical audiences
  • Experience publishing or contributing to open‑source ML projects
  • Comfort working end‑to‑end from research to production

Skills

  • Python
  • PyTorch
  • Large Language Model (LLM) development
  • Model training & fine‑tuning
  • Distributed training (FSDP/DeepSpeed)
  • Model serving frameworks (HuggingFace, vLLM, SGLang)
  • Experiment design & evaluation