
Job Description
Ovii's Interpretation of the Role
Gradera is hiring a Data Engineer to design, build and operate scalable data pipelines that feed digital‑twin platforms, real‑time operational systems and AI/ML workloads. The role sits in the Data & Digital Twin Foundation team in Hyderabad and works closely with data architects, simulation engineers and ML teams.
Role Snapshot
- Build and maintain scalable data pipelines
- Implement real‑time Kafka ingestion
- Apply Delta Live Tables for data quality
- Govern data via Unity Catalog
- Collaborate with ML and simulation engineers
- Support production data systems
- Develop warehouse solutions with Terradata
Must-Have Requirements
- Databricks
- PySpark
- Delta Lake
- Apache Kafka
- Delta Live Tables
- Unity Catalog
- Terradata
- hands‑on data engineering
- production‑grade pipeline development
Nice-to-Have Signals
- Digital twin pipeline development
- Operational state modeling
- Physics‑informed ML feature engineering
- Industrial domain knowledge (Manufacturing, Logistics, Transportation)
- Experience with distributed multidisciplinary teams
- Familiarity with time‑series data patterns
- Operational data modelling exposure
- digital twin platform pipelines
- industrial domain exposure
Work Setup
- Location: Hyderabad, India
- Work mode: ONSITE
- Employment type: Full-Time
Not Specified in JD
- Salary range
- Visa sponsorship
- Remote eligibility
- Education requirement
- Certifications
- Relocation
- Notice period
- Travel
- Security clearance
- Coding test
- Portfolio
- GitHub
- Writing sample
- Cover letter
What You'll Likely Work On
- Design, develop and maintain batch and streaming pipelines on Databricks using PySpark and Delta Lake
- Create high‑throughput Kafka ingestion pipelines for operational systems
- Build declarative pipelines with Delta Live Tables and enforce quality expectations
- Manage data lineage, access control and metadata through Unity Catalog
- Optimize pipeline performance, reliability and cost
- Produce clean, well‑documented code and participate in code reviews
- Partner with ML engineers to deliver feature‑engineered datasets for AI/ML models
- Support production data platforms through monitoring, troubleshooting and incident resolution
- Implement business‑intelligence warehouse solutions using Terradata
Good Fit If You Have
- Experience building pipelines for digital‑twin or simulation platforms
- Familiarity with time‑series data patterns and operational data modeling
- Exposure to manufacturing, logistics or transportation domains
- Proven ability to work in agile, cross‑functional teams
- Interest in physics‑informed or time‑series ML feature engineering
Skills
- Databricks
- PySpark
- Delta Lake
- Apache Kafka
- Delta Live Tables
- Unity Catalog
- Terradata
- Data pipeline engineering
- Real‑time data processing
- Data governance
- Data quality validation