- Location
- Hungary (Debrecen)
- Workplace
- Remote
- Type
- Full-time
- Department
- Engineering
- Education
- Bachelor
- Source
- Workday
Description
***This role is for pipeline purposes only; while we don’t have an immediate opening, we frequently launch new positions, and relevant candidates will be contacted once an opportunity becomes available.***
The Data Governance DataHub Engineer is responsible for deploying, managing, and extending the enterprise metadata management and data governance platform built on LinkedIn's open-source DataHub. This role ensures the discoverability, lineage, ownership, and compliance of all data assets across the lakehouse ecosystem.
KEY RESPONSIBILITIES
▸ Architect, deploy, and operate the DataHub metadata platform in cloud and on-premises environments
▸ Integrate DataHub with lakehouse sources including Databricks, Snowflake, Kafka, Airflow, and BI tools
▸ Configure and maintain automated metadata ingestion pipelines (DataHub Managed Ingestion)
▸ Build and maintain data lineage graphs across the full data stack (source to dashboard)
▸ Implement data classification, tagging, and business glossary within DataHub
▸ Enable and enforce data ownership assignment and stewardship workflows
▸ Collaborate with data governance team to translate policies into platform controls
▸ Develop custom DataHub plugins and APIs to extend platform capabilities
▸ Monitor DataHub platform health, storage, and ingestion job performance
▸ Train data stewards, owners, and consumers on DataHub usage and governance standards
REQUIRED QUALIFICATIONS
▸ 5+ years in data engineering or platform engineering roles
▸ 2+ years of hands-on experience with DataHub deployment and configuration
▸ Strong proficiency in Python for writing custom DataHub recipes and plugins
▸ Experience with Docker, Kubernetes, and Helm chart management
▸ Familiarity with metadata standards: OpenLineage, OpenAPI, Apache Atlas metadata models
▸ Understanding of data governance concepts: data lineage, classification, stewardship, and RBAC
▸ Experience integrating with REST APIs and event-driven architectures (Kafka)
▸ Bachelor's degree in Computer Science or related technical field
PREFERRED QUALIFICATIONS
▸ Experience with Atlan, Alation, or Collibra as complementary or alternative catalog tools
▸ Familiarity with GDPR/CCPA data inventory and lineage reporting requirements
▸ Knowledge of great expectations or Monte Carlo for data observability
At Dynata, we are committed to fostering an inclusive, accessible environment, where all employees and customers feel valued, respected and supported. We are dedicated to building a workforce that reflects the diversity of our customers and communities in which we live and serve. Dynata welcomes and encourages applications from people with disabilities. We are committed to an inclusive work culture for all our employees. Accommodations by request can be made for all aspects of the selection process.
#LI-DI1
#LI-Remote