Jobless Developer

Data Governance / DataHub Engineer

DebrecenRemoteFull-time

AI Summary

Deploys, manages, and extends the enterprise DataHub metadata and data governance platform to ensure data discoverability, lineage, ownership, and compliance across the lakehouse ecosystem.

About this role

The Data Governance DataHub Engineer is responsible for deploying, managing, and extending the enterprise metadata management and data governance platform built on LinkedIn's open-source DataHub. This role ensures the discoverability, lineage, ownership, and compliance of all data assets across the lakehouse ecosystem. 

KEY RESPONSIBILITIES

▸  Architect, deploy, and operate the DataHub metadata platform in cloud and on-premises environments 

▸  IntegrateDataHub with lakehouse sources including Databricks, Snowflake, Kafka, Airflow, and BI tools 

▸  Configure and maintain automated metadata ingestion pipelines (DataHub Managed Ingestion) 

▸  Build and maintain data lineage graphs across the full data stack (source to dashboard) 

▸  Implement data classification, tagging, and business glossary within DataHub

▸  Enable and enforce data ownership assignment and stewardship workflows 

▸  Collaborate with data governance team to translate policies into platform controls 

▸  Develop custom DataHub plugins and APIs to extend platform capabilities 

▸  MonitorDataHub platform health, storage, and ingestion job performance 

▸  Train data stewards, owners, and consumers on DataHub usage and governance standards 

REQUIRED QUALIFICATIONS

▸  5+ years in data engineering or platform engineering roles 

▸  2+ years of hands-on experience with DataHub deployment and configuration 

▸  Strongproficiency in Python for writing custom DataHub recipes and plugins 

▸  Experience with Docker, Kubernetes, and Helm chart management 

▸  Familiarity with metadata standards: OpenLineage, OpenAPI, Apache Atlas metadata models 

▸  Understanding of data governance concepts: data lineage, classification, stewardship, and RBAC 

▸  Experience integrating with REST APIs and event-driven architectures (Kafka) 

▸  Bachelor's degree in Computer Science or related technical field 

PREFERRED QUALIFICATIONS

▸  Experience with Atlan, Alation, or Collibra as complementary or alternative catalog tools 

▸  Familiarity with GDPR/CCPA data inventory and lineage reporting requirements 

▸  Knowledge of great expectations or Monte Carlo for data observability 

At Dynata, we are committed to fostering an inclusive, accessible environment, where all employees and customers feel valued, respected and supported. We are dedicated to building a workforce that reflects the diversity of our customers and communities in which we live and serve. Dynata welcomes and encourages applications from people with disabilities. We are committed to an inclusive work culture for all our employees. Accommodations by request can be made for all aspects of the selection process.

#LI-DI1

#LI-Remote

Skills

AirflowApache AtlasCCPADatabricksDataHubDockerGDPRGreat ExpectationsHelmKafkaKubernetesMonte CarloOpenAPIOpenLineagePythonRBACREST APISnowflake

Explore related jobs

Browse these categories

Market data for this role

All reports →