Jr Data Engineer
Job Description
About this Position As a Junior Data Engineer on the Customer Intelligence team at Henkel, you will help us build, scale, and maintain our global Customer Data Platform (CDP). You will be a key player in creating a 360-degree view of our customers, transforming raw data into actionable insights. Working alongside experienced engineers and architects, this role is perfect for a growth-minded developer with strong foundational coding skills who is eager to dive deep into modern cloud data environments, data mesh concepts, and
end-to-end data pipeline development.
What You´ll Do
end-to-end data pipeline development.
What You´ll Do
- Pipeline Development (35%): Collaborate on building and maintaining foundational end-to-end (e2e) ETL/ELT pipelines to extract data from internal and external sources (APIs, Service Connectors, Delta Sharing, and Kafka streams).
- Data Reshaping & Modeling (25%): Use PySpark and SQL within Databricks to clean, transform, and reshape raw ingestion data. Gain hands-on exposure to our Data Vault customer and contact modeling frameworks.
- Activation & Insights (10%): Support the development of our Gold/Business layer views to power business dashboards and assist in feeding processed data back into the Adobe Experience Platform (AEP) via APIs and streaming pipelines.
- Orchestration & DevOps (15%): Write clean, version-controlled code using Git, participate in CI/CD DevOps workflows, and assist in monitoring pipelines using Databricks Workflows and Delta Live Tables (DLT).
- Quality & Governance (15%): Write automated data quality tests, maintain clear technical documentation, and ensure data governance standards are strictly followed across your daily tasks.
- Solid programming skills in Python and high proficiency in writing and optimizing SQL queries.
- Working with third-party APIs
- Foundational understanding of core data warehousing concepts and ETL/ELT processes.
- Basic familiarity with Git for version control.
- Good communication skills with the ability to explain technical ideas clearly and work collaboratively within an agile team environment.
- Hands-on exposure to Databricks, PySpark, or Delta Live Tables (DLT) is a plus.
- Basic understanding of automated CI/CD DevOps concepts.
- Exposure to streaming data concepts (e.g., Kafka).
- Vision-Driven: You care about the why behind the code and the broader goal, rather than just executing isolated, task-driven assignments.
- Growth & Curiosity: A proactive learner who is excited to pick up advanced concepts like Data Vault modeling and decentralized Data Mesh paradigms.
- Flexible work scheme with flexible hours, hybrid work model, and work from anywhere policy for up to 30 days per year
- Diverse national and international growth opportunities
- Global wellbeing standards with health and preventive care programs
- Gender-neutral parental leave for a minimum of 8 weeks
- Employee Share Plan with voluntary investment and Henkel matching shares
- Comprehensive Health Insurance for employee + dependents
- Employee Assistance Programme provides a wide range of mental health and wellbeing benefits
