Ovii Job Board

Senior Data Engineer

Skan AI

Bengaluru, India • Onsite - Bengaluru, India • Full-Time • 5+ years

Posted 2026-07-01 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

As a Senior Data Engineer at Skan AI, you will design, build, and operate real‑time Flink pipelines that sync PostgreSQL data into StarRocks, while mentoring engineers and driving best‑practice adoption across the team.

Role Snapshot

  • Design and implement high‑performance Flink ETL/ELT pipelines
  • Lead data‑engineering initiatives and mentor junior engineers
  • Optimize pipeline throughput, latency, and resource usage
  • Deploy and manage Flink clusters on Kubernetes/GKE
  • Collaborate with analysts, infra, and DBA teams
  • Own reliability, monitoring, and incident response

Must-Have Requirements

  • Apache Flink
  • Java/Scala/Python
  • PostgreSQL CDC
  • StarRocks
  • Kafka
  • Distributed systems concepts
  • Kubernetes/GKE
  • Helm
  • Harness CI/CD
  • Prometheus & Grafana
  • data engineering
  • real‑time streaming pipelines
  • Apache Flink production experience

Nice-to-Have Signals

  • StarRocks Flink Connector
  • Flink on Kubernetes
  • Iceberg / Delta Lake / Hudi
  • Apache Airflow or Prefect
  • Advanced observability integrations
  • Data mining / large‑scale data processing domain experience
  • Data governance, lineage, metadata management
  • StarRocks connector development
  • Kubernetes‑based Flink deployments
  • CI/CD for data workloads

Work Setup

  • Location: Bengaluru, India
  • Work mode: ONSITE
  • Employment type: Full-Time

Eligibility Gates

  • Visa sponsorship: unknown

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility
  • Education requirement
  • Certifications
  • Relocation
  • Notice period
  • Travel
  • Security clearance
  • Coding test
  • Portfolio
  • GitHub
  • Writing sample
  • Cover letter

What You'll Likely Work On

  • Build and maintain real‑time Flink pipelines for PostgreSQL‑to‑StarRocks sync
  • Architect CDC solutions using Flink CDC connectors and Debezium
  • Write transformation logic with Flink SQL and DataStream APIs
  • Develop custom Flink connectors and StarRocks sinks
  • Tune performance (parallelism, memory, back‑pressure) and optimize StarRocks loading
  • Implement fault‑tolerant, exactly‑once pipelines with checkpointing and state backends
  • Set up monitoring, alerts, and data‑quality checks using Prometheus/Grafana
  • Lead a small team of data engineers, conduct code reviews, and drive delivery

Good Fit If You Have

  • Hands‑on experience with StarRocks Flink Connector
  • Flink on Kubernetes (Operator or native mode)
  • Familiarity with Harness CI/CD and Helm chart management
  • Knowledge of lakehouse formats such as Iceberg, Delta Lake, or Hudi
  • Prior technical‑lead or senior engineering role with delivery accountability

Skills

  • Apache Flink (SQL & DataStream API)
  • Java / Scala / Python
  • PostgreSQL CDC (replication slots, WAL)
  • StarRocks (columnar OLAP)
  • Kafka (messaging)
  • Distributed systems concepts (fault tolerance, exactly‑once, state management)
  • Kubernetes / GKE deployment
  • Helm & Harness CI/CD
  • Prometheus & Grafana observability
  • Leadership & mentoring