Ovii Job Board

Senior Artificial Intelligence Researcher

Innovaccer

San Francisco, United States • Onsite - San Francisco, United States • Full-Time

Posted 2026-08-24 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

We seek a senior AI researcher to design, build, and ship production‑grade AI solutions—including LLMs, agents, and retrieval‑augmented generation—within a healthcare data platform. The role blends hands‑on model engineering, experimental rigor, and cross‑functional collaboration with product, engineering, and data teams.

Role Snapshot

  • Senior AI researcher
  • Production‑grade AI systems
  • LLM, RAG, and AI agent development
  • Cross‑functional collaboration
  • Python & PyTorch engineering
  • Large‑scale model training & deployment

Must-Have Requirements

  • Python
  • PyTorch
  • Multi‑node GPU training
  • Model fine‑tuning
  • Production model deployment
  • Data pipeline ownership
  • Research‑to‑production AI system development
  • Large‑scale model training and serving
  • Data pipeline engineering for model training
  • MS or PhD in Computer Science, Machine Learning, or related quantitative field

Nice-to-Have Signals

  • Mentoring or technical direction
  • Research publications at top conferences
  • Open‑source ML contributions
  • Familiarity with HuggingFace, DeepSpeed/FSDP, vLLM/SGLang
  • Top‑conference publications
  • Open‑source contributions
  • Mentoring junior engineers

Work Setup

  • Location: San Francisco, United States
  • Work mode: ONSITE
  • Employment type: Full-Time

Eligibility Gates

  • Visa sponsorship: unknown

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility
  • Coding test

What You'll Likely Work On

  • Design, prototype, and ship AI‑powered features such as LLM‑based agents and RAG pipelines
  • Build and scale the model layer, selecting model sizes and composing multiple models into a product‑ready system
  • Run rigorous experiments, hypothesis testing, and evaluation to validate model performance
  • Optimize training and serving pipelines on multi‑GPU clusters using DeepSpeed/FSDP or equivalent
  • Own end‑to‑end data pipelines for training, including deduplication, filtering, and format normalization
  • Collaborate with product, engineering, and data teams to translate research ideas into production outcomes
  • Mentor junior engineers and help set technical direction as the team grows (preferred)

Good Fit If You Have

  • Hands‑on curiosity and ability to move research from paper to production
  • Experience shipping models that meet real‑world accuracy and latency targets
  • Strong communication skills for explaining AI results to clinicians and non‑technical stakeholders
  • Track record of publications or impactful open‑source contributions

Skills

  • Python programming
  • PyTorch
  • Large language model fine‑tuning (parameter‑efficient & full)
  • Multi‑node GPU training (tens of GPUs, >10B‑parameter models)
  • Model serving stacks (HuggingFace, DeepSpeed/FSDP, vLLM/SGLang)
  • Data pipeline engineering for training data
  • Experiment design & evaluation
  • Research publication record (NeurIPS, ICML, ICLR, ACL, EMNLP)
  • Open‑source ML contributions
  • Mentoring or technical direction for engineers