Ovii Job Board

Consultant- Data Aggregation(336)

beghou

Pune, India • Hybrid - Pune, India • Full-Time • 4+ years

Posted 2026-06-11 Tech & Engg

Apply on employer site

Job Description

Ovii's Interpretation of the Role

The Consultant – Data Aggregation will secure, clean, and combine patient‑level data from multiple health‑care sources, applying tokenization and de‑identification techniques while ensuring HIPAA and HISEC compliance. The role blends consulting delivery with hands‑on Python/PySpark development and project coordination for U.S. pharmaceutical clients.

Role Snapshot

  • Secure patient data aggregation
  • HIPAA & HISEC compliance
  • Python & PySpark development
  • Data pipeline design & management
  • Consulting project delivery
  • Cross‑functional stakeholder coordination

Must-Have Requirements

  • Python
  • PySpark
  • Data tokenization / de‑identification
  • HIPAA / GDPR compliance
  • Data pipeline management
  • consulting experience
  • data management and analytics
  • U.S. pharmaceutical datasets
  • Engineering/Master’s Degree in Computer Science or related field

Nice-to-Have Signals

  • Snowflake or Redshift
  • Airflow or Databricks workflows
  • Microsoft Office products
  • experience with Snowflake, Redshift, Airflow, Databricks

Work Setup

  • Location: Pune, India
  • Work mode: HYBRID
  • Remote scope: UNSPECIFIED
  • Employment type: Full-Time

Eligibility Gates

  • Visa sponsorship: unknown

Not Specified in JD

  • Visa sponsorship
  • Salary range
  • Remote eligibility
  • Certifications
  • Relocation
  • Notice period
  • Travel
  • Security clearance
  • Coding test
  • Portfolio
  • GitHub
  • Cover letter

What You'll Likely Work On

  • Collect, clean, and aggregate patient‑level data from EHRs, labs, and external databases.
  • Apply tokenization and expert‑determination methods to de‑identify data while preserving linkage capability.
  • Build and maintain secure data pipelines and repositories in a HISEC‑compliant environment.
  • Create SOPs and documentation for secure data handling and support audit or governance reviews.
  • Develop end‑to‑end aggregation solutions using Python and PySpark, integrating with Snowflake/Redshift and workflow tools.
  • Manage multiple client projects, lead meetings, and ensure timely delivery with minimal supervision.

Good Fit If You Have

  • Enjoys collaborating with cross‑functional and international teams.
  • Comfortable working with U.S. pharmaceutical datasets and regulatory frameworks.
  • Can manage several projects simultaneously and meet deadlines.
  • Willing to provide 3‑4 hour overlap with U.S. working hours.

Skills

  • Python
  • PySpark
  • Data tokenization / de‑identification
  • HIPAA / GDPR compliance
  • Data modeling & analytics
  • Snowflake or Redshift
  • Airflow or Databricks workflows
  • Data quality monitoring