Ovii Job Board

Principal AI Researcher

Innovaccer

San Francisco, United States • Onsite - San Francisco, United States • Full-Time

Posted 2026-08-24 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

Innovaccer seeks a Principal AI Researcher to design, develop, and ship production‑grade AI systems for healthcare. The role spans the full AI lifecycle—from research prototypes to large‑scale deployed models—working closely with product, engineering, and data teams.

Role Snapshot

  • Lead end‑to‑end AI product development
  • Design and ship production LLM solutions
  • Build large‑scale distributed training pipelines
  • Mentor engineers and set technical direction
  • Collaborate across product, data, and engineering
  • Drive research‑to‑production translation

Must-Have Requirements

  • Python
  • PyTorch
  • Large‑scale distributed training frameworks (DeepSpeed, FSDP, etc.)
  • Model serving stacks (HuggingFace, vLLM, SGLang)
  • Multi‑GPU training (tens of GPUs)
  • Experiment design & evaluation
  • Data pipeline engineering
  • First‑author publications at top ML conferences
  • Research publications at top conferences
  • Production deployment of large AI models
  • Large‑scale distributed training
  • MS in Computer Science, Machine Learning, or related quantitative field
  • PhD in Computer Science, Machine Learning, or related quantitative field

Nice-to-Have Signals

  • Mentoring engineers
  • Setting technical direction for AI team
  • Familiarity with training/serving stack (HuggingFace, FSDP, DeepSpeed, vLLM, SGLang)
  • Exposure to multi‑GPU training at lab scale
  • Technical leadership

Work Setup

  • Location: San Francisco, United States
  • Work mode: ONSITE
  • Employment type: Full-Time

What You'll Likely Work On

  • Prototype and productionize LLM‑based applications, including retrieval‑augmented generation and AI agents
  • Design, train, and fine‑tune large models (10B+ parameters) on multi‑node GPU clusters
  • Build and maintain data pipelines for training data deduplication, filtering, and format normalization
  • Develop model serving infrastructure and meet product accuracy and latency targets
  • Run hypothesis‑driven experiments, ablations, and rigorous evaluations
  • Explain model behavior to clinicians, operators, and other non‑technical stakeholders
  • Define the technical roadmap for in‑house modeling and guide the AI team
  • Mentor junior engineers and help set team direction (preferred)

Good Fit If You Have

  • First‑author papers at top ML conferences (NeurIPS, ICML, ICLR, ACL, EMNLP)
  • Hands‑on experience shipping models that meet production accuracy and latency goals
  • Ability to translate complex AI concepts for non‑technical audiences
  • Experience mentoring engineers or setting technical direction (preferred)

Skills

  • Python
  • PyTorch
  • Large‑scale distributed training (DeepSpeed/FSDP)
  • LLM fine‑tuning & parameter‑efficient methods
  • Model serving stacks (HuggingFace, vLLM, SGLang)
  • Experiment design & evaluation
  • Data pipeline engineering
  • Research publication (NeurIPS, ICML, etc.)
  • Multi‑GPU training (tens of GPUs)
  • Production performance optimization