Senior Data Engineer
Copenhagen
Contract
Negotiable
Medical Affairs
CR/612869_1791184935
Senior Data Engineer
Key Responsibilities
- Design, develop, and optimize data solutions using Azure Databricks, Python/PySpark, Spark SQL, and Delta Lake.
- Develop and maintain data pipelines following the Medallion Architecture (Bronze, Silver, Gold) to deliver scalable and trusted data products.
- Implement ingestion frameworks for structured and semi-structured data from APIs, JSON files, databases, and other enterprise sources.
- Manage and optimize data storage solutions using Azure Data Lake Storage Gen2 (ADLS Gen2) and associated Azure services.
- Build, schedule, and support workloads using Databricks Workflows/Jobs, with experience in Delta Live Tables (DLT) or Lakeflow considered a strong advantage.
- Implement schema management practices, including schema enforcement and schema evolution.
- Develop and maintain data quality controls, validation frameworks, reconciliation processes, and automated testing.
- Design and implement data transformation, cleansing, and standardization processes to support business reporting and analytics.
- Establish metadata management, data lineage, and governance capabilities using Unity Catalog and related technologies.
- Develop API-based ingestion and export integrations for upstream and downstream systems.
- Support solutions involving document metadata management and references to PDF and eLabel assets.
- Drive performance tuning, partitioning strategies, and platform optimization to improve scalability and efficiency.
- Implement CI/CD practices for notebooks, code deployments, and automated testing throughout the development lifecycle.
- Develop monitoring, alerting, error-handling, and reprocessing capabilities to ensure operational stability and reliability.
- Support regulated data environments with a strong focus on traceability, auditability, reproducibility, and controlled release management.
