
Job Description
Ovii's Interpretation of the Role
As a Senior Data Engineer at Skan AI, you will design, build, and operate real‑time Flink pipelines that sync PostgreSQL data into StarRocks, while mentoring engineers and driving best‑practice adoption across the team.
Role Snapshot
- Design and implement high‑performance Flink ETL/ELT pipelines
- Lead data‑engineering initiatives and mentor junior engineers
- Optimize pipeline throughput, latency, and resource usage
- Deploy and manage Flink clusters on Kubernetes/GKE
- Collaborate with analysts, infra, and DBA teams
- Own reliability, monitoring, and incident response
Must-Have Requirements
- Apache Flink
- Java/Scala/Python
- PostgreSQL CDC
- StarRocks
- Kafka
- Distributed systems concepts
- Kubernetes/GKE
- Helm
- Harness CI/CD
- Prometheus & Grafana
- data engineering
- real‑time streaming pipelines
- Apache Flink production experience
Nice-to-Have Signals
- StarRocks Flink Connector
- Flink on Kubernetes
- Iceberg / Delta Lake / Hudi
- Apache Airflow or Prefect
- Advanced observability integrations
- Data mining / large‑scale data processing domain experience
- Data governance, lineage, metadata management
- StarRocks connector development
- Kubernetes‑based Flink deployments
- CI/CD for data workloads
Work Setup
- Location: Bengaluru, India
- Work mode: ONSITE
- Employment type: Full-Time
Eligibility Gates
- Visa sponsorship: unknown
Not Specified in JD
- Visa sponsorship
- Salary range
- Remote eligibility
- Education requirement
- Certifications
- Relocation
- Notice period
- Travel
- Security clearance
- Coding test
- Portfolio
- GitHub
- Writing sample
- Cover letter
What You'll Likely Work On
- Build and maintain real‑time Flink pipelines for PostgreSQL‑to‑StarRocks sync
- Architect CDC solutions using Flink CDC connectors and Debezium
- Write transformation logic with Flink SQL and DataStream APIs
- Develop custom Flink connectors and StarRocks sinks
- Tune performance (parallelism, memory, back‑pressure) and optimize StarRocks loading
- Implement fault‑tolerant, exactly‑once pipelines with checkpointing and state backends
- Set up monitoring, alerts, and data‑quality checks using Prometheus/Grafana
- Lead a small team of data engineers, conduct code reviews, and drive delivery
Good Fit If You Have
- Hands‑on experience with StarRocks Flink Connector
- Flink on Kubernetes (Operator or native mode)
- Familiarity with Harness CI/CD and Helm chart management
- Knowledge of lakehouse formats such as Iceberg, Delta Lake, or Hudi
- Prior technical‑lead or senior engineering role with delivery accountability
Skills
- Apache Flink (SQL & DataStream API)
- Java / Scala / Python
- PostgreSQL CDC (replication slots, WAL)
- StarRocks (columnar OLAP)
- Kafka (messaging)
- Distributed systems concepts (fault tolerance, exactly‑once, state management)
- Kubernetes / GKE deployment
- Helm & Harness CI/CD
- Prometheus & Grafana observability
- Leadership & mentoring