
Job Description
Ovii's Interpretation of the Role
Gradera seeks a Data Engineer to design, build, and maintain scalable data pipelines that feed digital‑twin platforms, real‑time operational systems, and AI/ML workloads. The role works closely with data architects, simulation engineers, and ML teams to deliver high‑quality, governed datasets for intelligent decision‑making.
Role Snapshot
- Design and maintain scalable data pipelines
- Build real‑time and batch ingestion with Kafka
- Implement data quality via Delta Live Tables
- Ensure governance with Unity Catalog
- Collaborate with ML and simulation engineers
- Optimize performance, reliability, and cost
- Support production systems and troubleshoot incidents
- Develop business data warehouse solutions using Terradata
Nice-to-Have Signals
- 7+ years hands‑on data engineering experience
- Delta Live Tables for declarative pipeline development
- Agile, cross‑functional team experience
- Familiarity with time‑series data patterns
- Operational state modeling for real‑time systems
- Physics‑informed ML feature engineering exposure
- Experience with digital‑twin or simulation pipelines
- Exposure to manufacturing, logistics, or transportation domains
- production‑grade data pipelines
- Delta Live Tables development
- agile cross‑functional collaboration
- Manufacturing
- Logistics
- Transportation
Work Setup
- Location: Hyderabad, India
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Education requirement
- Certifications
- Relocation
- Travel
- Security clearance
- Coding test
- Portfolio
- GitHub
- Writing sample
- Cover letter
What You'll Likely Work On
- Design, develop, and maintain scalable pipelines on Databricks using PySpark and Delta Lake
- Create real‑time and batch ingestion pipelines from operational systems via Apache Kafka
- Apply data transformations for digital‑twin platforms and operational analytics
- Integrate Kafka event streams with Databricks for live state updates
- Enforce data quality and validation with Delta Live Tables expectations
- Implement governance, lineage, and access control through Unity Catalog
- Optimize pipeline performance, reliability, and cost efficiency
- Collaborate with ML engineers to deliver feature‑engineered datasets
- Support production data systems with monitoring, troubleshooting, and incident resolution
- Build business‑intelligence data warehouse solutions using Terradata
Good Fit If You Have
- Enjoys building production‑grade data pipelines at scale
- Comfortable with real‑time event streaming and Kafka
- Familiar with time‑series data and operational modeling
- Thrives in agile, cross‑functional team environments
- Interest in digital‑twin or simulation platforms
Skills
- Databricks (Delta Lake, Unity Catalog)
- PySpark
- Apache Kafka
- Delta Live Tables
- Terradata
- SQL / Databricks SQL
- Data modeling & time‑series patterns
- Agile cross‑functional teamwork