
Lead Data & AI Engineer
comprinno
Posted 2026-05-08
INR 1,800,000 - INR 2,700,000 per year
Tech & Engg
Job Description
Ovii's Interpretation of the Role
The Lead Data & AI Engineer will architect and deliver enterprise‑scale data platforms, machine‑learning pipelines, and generative‑AI solutions on AWS. The role blends deep technical expertise with customer‑facing consulting, mentorship, and practice‑building responsibilities.
Role Snapshot
- Lead design of scalable data & AI platforms
- Build production ML and GenAI applications
- Implement end‑to‑end MLOps pipelines
- Architect AWS‑native data & AI services
- Mentor and grow Data & AI engineers
- Drive customer workshops and presales
- Create reusable frameworks and accelerators
Must-Have Requirements
- Python
- SQL
- Apache Spark
- Kafka
- Airflow
- AWS SageMaker
- AWS Bedrock
- AWS Glue
- AWS Athena
- AWS Redshift
- AWS EMR
- AWS Kinesis
- AWS Lake Formation
- AWS OpenSearch
- TensorFlow
- PyTorch
- Scikit‑Learn
- MLflow
- Kubeflow
- SageMaker Pipelines
- LangChain
- Hugging Face
- OpenAI APIs
- Anthropic APIs
- Pinecone
- Weaviate
- Chroma
- Data Engineering
- Machine Learning
- Generative AI
- AWS cloud services
- Bachelor's or Master's in Computer Science
- Bachelor's or Master's in Data Engineering
- Bachelor's or Master's in Data Science
- Bachelor's or Master's in Artificial Intelligence
- Bachelor's or Master's in Statistics
- Bachelor's or Master's in Mathematics
- AWS Certified Data Analytics – Specialty
- AWS Machine Learning – Specialty
- AWS Solutions Architect – Professional
- Must hold one of: AWS Certified Data Analytics – Specialty, AWS Machine Learning – Specialty, AWS Solutions Architect – Professional
Nice-to-Have Signals
- Databricks
- Snowflake
- BigQuery
- Databricks Data Engineer Professional
- SnowPro Certification
- Google Professional Data Engineer
- Azure Data Engineer Associate
- AWS AI Practitioner
- Data governance platforms (familiarity)
- Regulated industry experience (optional)
- Regulated industry experience
- Open‑source contributions
Work Setup
- Location: Pune, India
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Coding test
What You'll Likely Work On
- Design and implement data lakes, lakehouses, warehouses, and real‑time streaming architectures on AWS
- Develop, train, and operationalize ML models and generative‑AI solutions using LLMs and RAG pipelines
- Build and maintain MLOps frameworks for model lifecycle, monitoring, and governance
- Architect end‑to‑end cloud‑native data & AI services leveraging SageMaker, Bedrock, Glue, Athena, and related AWS services
- Lead discovery workshops, translate business needs into technical designs, and support presales activities
- Mentor engineers, conduct architecture and code reviews, and define practice standards
- Contribute reusable accelerators, frameworks, and intellectual property for the Data & AI practice
Good Fit If You Have
- Experience in regulated sectors such as BFSI, healthcare, manufacturing, or telecom
- Open‑source contributions or community involvement in data/AI ecosystems
- Familiarity with data‑governance and observability platforms
- Strong consulting mindset and ability to drive business outcomes
Skills
- Python
- SQL
- Apache Spark
- Kafka
- Airflow
- AWS SageMaker / Bedrock
- AWS Glue / Athena / Redshift / EMR / Kinesis / Lake Formation / OpenSearch
- TensorFlow / PyTorch / Scikit‑Learn
- MLflow / Kubeflow / SageMaker Pipelines
- LangChain / Hugging Face / OpenAI / Anthropic
- Vector DBs (Pinecone, Weaviate, Chroma)
- Data Mesh & Lakehouse concepts