
Job Description
Ovii's Interpretation of the Role
Innovaccer seeks an AI Engineer to design, develop, and ship production‑grade AI solutions for its healthcare platform. You will partner with product, engineering, and data teams to build LLM‑based applications, AI agents, and retrieval‑augmented generation pipelines from research to deployment.
Role Snapshot
- Build production‑grade AI models and pipelines
- Develop LLM‑based applications and AI agents
- Deploy and optimize AI services at scale
- Collaborate across product, engineering, and data teams
- Design experiments and evaluate model performance
- Write production‑ready Python/PyTorch code
Must-Have Requirements
- Python
- PyTorch
- AI model development
- LLM solutions
- Production deployment
- MS in Computer Science, Machine Learning, or related quantitative field
- PhD in Computer Science, Machine Learning, or related quantitative field
Nice-to-Have Signals
- Familiarity with HuggingFace, DeepSpeed/FSDP, vLLM, SGLang
- Multi‑GPU training experience
- Knowledge of sharding strategies
- Open‑source ML contributions
- First‑author publications at top conferences
- Some exposure to multi‑GPU training
- Research publications
- Open‑source contributions
- BS with substantial research or open‑source work (exceptional candidates)
Work Setup
- Location: San Francisco, United States
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Certification requirements
- Shift or travel expectations
What You'll Likely Work On
- Design, prototype, and ship AI models and pipelines that power healthcare applications
- Implement LLM‑based solutions, AI agents, and retrieval‑augmented generation workflows
- Fine‑tune open‑weight models, choose parameter‑efficient or full fine‑tuning, and meet product‑level accuracy targets
- Build and maintain training and serving infrastructure using HuggingFace, DeepSpeed/FSDP, vLLM/SGLang, or equivalents
- Collaborate with product, engineering, and data teams to translate research ideas into production features
- Communicate model behavior and results to clinicians and other non‑technical stakeholders
Good Fit If You Have
- Strong curiosity for cutting‑edge AI research
- Ability to explain technical concepts to non‑technical audiences
- Experience publishing or contributing to open‑source ML projects
- Comfort working end‑to‑end from research to production
Skills
- Python
- PyTorch
- Large Language Model (LLM) development
- Model training & fine‑tuning
- Distributed training (FSDP/DeepSpeed)
- Model serving frameworks (HuggingFace, vLLM, SGLang)
- Experiment design & evaluation