
Job Description
Ovii's Interpretation of the Role
Senior Data Engineer role based in Chennai, building end‑to‑end data pipelines that crawl public brand data, store it cleanly, and enrich it with sentiment and topic signals for analyst use.
Role Snapshot
- Senior‑level data engineering
- End‑to‑end pipeline construction
- Web crawling across multiple sources
- Data quality & incremental processing
- Sentiment & topic enrichment
- Design documentation & evaluation
Must-Have Requirements
- Data engineering
- Web crawling / scraping
- ETL pipeline development
- Data storage & schema design
- Sentiment analysis
- Topic/theme extraction
- Version control (Git)
- Design documentation
- Web crawling
- Sentiment enrichment
Work Setup
- Location: Chennai, India
- Work mode: ONSITE
- Employment type: Full-Time
Application Requirements
- GitHub required
- Coding test mentioned
- Assessment mentioned
Not Specified in JD
- Salary range
- Visa sponsorship
- Remote eligibility
- Education requirement
What You'll Likely Work On
- Crawl brand‑related public data from at least five distinct sources, handling pagination, rate limits, and retries.
- Create an incremental, idempotent ETL pipeline that unifies the crawled data into a clean, queryable store with deduplication.
- Implement robust data‑quality checks, logging, and metrics to support daily unattended operation.
- Enrich the unified data with sentiment scores and thematic topics, and evaluate accuracy against hand‑labeled samples.
- Expose an analytical query that answers a real business question about the brand using the stored data.
- Write a concise design document covering architecture, source choices, scaling, failure recovery, and cost considerations.
- Deliver a Git repository with README, code, sample data, and a short walkthrough video.
Good Fit If You Have
- Enjoys making design decisions and justifying technology choices.
- Comfortable documenting reasoning and trade‑offs in a design doc.
- Able to build reliable, repeatable data pipelines that run unattended.
Skills
- Data engineering
- Web crawling / scraping
- ETL pipeline development
- Data storage & schema design
- Sentiment analysis
- Topic/theme extraction
- Version control (Git)
- Design documentation