Dynata

Data Governance / DataHub Engineer

Dynata  •  Debrecen, HU (Remote)  •  4 days ago
Apply
AI can make mistakes so check important info. Chat history is never stored.

Job Description

***This role is for pipeline purposes only; while we don’t have an immediate opening, we frequently launch new positions, and relevant candidates will be contacted once an opportunity becomes available.***

The Data Governance DataHub Engineer is responsible for deploying, managing, and extending the enterprise metadata management and data governance platform built on LinkedIn's open-source DataHub This role ensures the discoverability, lineage, ownership, and compliance of all data assets across the lakehouse ecosystem. 

KEY RESPONSIBILITIES

  Architect, deploy, and operate the DataHub metadata platform in cloud and on-premises environments 

  IntegrateDataHub with lakehouse sources including Databricks, Snowflake, Kafka, Airflow, and BI tools 

  Configure and maintain automated metadata ingestion pipelines (DataHub Managed Ingestion) 

  Build and maintain data lineage graphs across the full data stack (source to dashboard) 

  Implement data classification, tagging, and business glossary within DataHub

  Enable and enforce data ownership assignment and stewardship workflows 

  Collaborate with data governance team to translate policies into platform controls 

  Develop custom DataHub plugins and APIs to extend platform capabilities 

  MonitorDataHub platform health, storage, and ingestion job performance 

  Train data stewards, owners, and consumers on DataHub usage and governance standards 

REQUIRED QUALIFICATIONS

  5+ years in data engineering or platform engineering roles 

  2+ years of hands-on experience with DataHub deployment and configuration 

  Strongproficiency in Python for writing custom DataHub recipes and plugins 

  Experience with Docker, Kubernetes, and Helm chart management 

  Familiarity with metadata standards: OpenLineage, OpenAPI, Apache Atlas metadata models 

  Understanding of data governance concepts: data lineage, classification, stewardship, and RBAC 

  Experience integrating with REST APIs and event-driven architectures (Kafka) 

  Bachelor's degree in Computer Science or related technical field 

PREFERRED QUALIFICATIONS

  Experience with Atlan, Alation, or Collibra as complementary or alternative catalog tools 

  Familiarity with GDPR/CCPA data inventory and lineage reporting requirements 

  Knowledge of great expectations or Monte Carlo for data observability 

At Dynata, we are committed to fostering an inclusive, accessible environment, where all employees and customers feel valued, respected and supported. We are dedicated to building a workforce that reflects the diversity of our customers and communities in which we live and serve. Dynata welcomes and encourages applications from people with disabilities. We are committed to an inclusive work culture for all our employees. Accommodations by request can be made for all aspects of the selection process.

#LI-DI1

#LI-Remote

Dynata

About Dynata

At Dynata, we deliver the highest quality first-party data to help businesses around the world gain precise insights, activate the right audiences, and confidently measure impact. With industry-leading responder accuracy, reliability, and commitment to continuous improvement, Dynata is the trusted foundation for smarter decision-making. We serve more than 6,000 market research, media and advertising agencies, publishers, consulting and investment firms, and corporate customers in North America, South America, Europe and the Asia-Pacific region.

Learn more at www.dynata.com. 

Industry
Research & Polling
Company Size
1,001-5,000 employees
Headquarters
Shelton, Connecticut
Year Founded
Unknown
Social Media