Unlock Full Resume Report

New offer - be the first one to apply!

September 3, 2026

Data Governance & Metadata Engineer

Mid • On-site

Krakow, MA, Poland

Key Responsibilities

Metadata Engineering & Automation

  • Model document data structures, metadata relationships, and end-to-end lineage across data platforms, including lakehouse architectures.
  • Configure and maintain metadata workflows, connectors, and governance assets.
  • Build automations and integrations in Java and Python; strong pandas expertise is required.
  • Integrate technical metadata from Databricks (Spark, Delta Lake, Unity Catalog) and Microsoft Purview into governance tools.
  • Operate and troubleshoot metadata services on Linux, including logs, services, and deployments.

Data Governance

  • Maintain data domains, dictionaries, glossaries, classification models, and stewardship structures.
  • Define and enforce metadata standards and data quality rules across analytical platforms.
  • Support impact analysis, remediation, and governance alignment for data products and lakehouse use cases.

Collibra / Purview

  • Maintain metadata, lineage, stewardship models, and workflows in Collibra.
  • Integrate Collibra with technical metadata sources, including Databricks, Unity Catalog, Microsoft Purview, and SQL engines, through APIs, scanners, or pipelines.
  • Align governance models between Collibra and Purview, including glossaries, classifications, and lineage where applicable.
  • Provide onboarding, training, and high-quality documentation.

Cross-Functional Work

  • Collaborate with data engineers, platform teams, and product owners to ensure consistent governance standards in Databricks, Unity Catalog, and Purview.
  • Assess risks, impacts, and compliance aspects in data-related projects.
  • Translate technical platform concepts, including Spark, lakehouse, catalogs, and semantic layers, into clear governance artefacts.

Required Skills

  • Hands-on experience with metadata platforms and governance tooling.
  • Strong understanding of data modelling, metadata architectures, lineage, and catalog concepts.
  • Proficiency in Java and Python, with mandatory pandas expertise.
  • Solid Linux skills.
  • Experience with REST APIs, SQL, and metadata extraction.
  • Practical experience with Databricks and Unity Catalog.
  • Familiarity with Microsoft Purview concepts, including scanning, classifications, and lineage.
  • Knowledge of data governance and data quality frameworks.
  • Strong documentation and communication skills.
  • Experience with Git and versioning workflows.

Nice to Have

  • Advanced hands-on experience with Databricks, including Spark, Delta Lake, Unity Catalog, and jobs.
  • Experience integrating Purview and Collibra in hybrid governance setups.
  • Azure services and orchestration tools experience.
  • Metadata scanning, MDM, or lineage tooling experience.
  • Collibra workflow development.
  • Understanding of data security, classification, and access control models.
  • Familiarity with industry lineage standards and open metadata approaches.

What We Offer

  • Hybrid working model with office attendance 8 days a month and remote-working options.
  • Competitive salary and comprehensive benefits.
  • Learning and development environment with an emphasis on knowledge sharing and training.
  • International, collaborative working environment.
  • Inclusive workplace that values diversity and considers qualified applicants regardless of personal characteristics.
  • Accommodation can be requested during the application process for disability or other needs.

Similar jobs you might like