Data Engineer

Data

Data Engineer

Data
-, Warszawa +4 Locations

Haddad Brands

Full-time
B2B
Mid
Remote
3 767,29 - 6 458,21 USDNet per month - B2B

Job description

Along with our dynamic growth, we are seeking a self-motivated Data Engineer who thrives in a team environment. This is a long-term contract role focused on designing and building a robust, governed data platform on top of our ERP data — moving the organization from manual, expert-dependent reporting toward a modern, self-service architecture in Microsoft Fabric.

About Haddad Brands

Haddad Brands is a leading provider of children’s apparel and accessories for iconic global brands such as Nike, Levi’s, and Jordan. Founded in 1948 by the Haddad brothers, we remain a privately held family business with over 100 years of industry experience, dedicated to delivering high-quality products and innovative solutions.

Our data spans the full value chain — production, sourcing, warehouse, wholesale and retail — and we are building a central, governed data platform in Microsoft Fabric to replace fragmented ERP reporting with reporting-ready and AI-ready datasets.

The Role

You will work directly with the data platform owner on building and maintaining the Lakehouse (Bronze → Silver → Gold), ingestion pipelines from AS400 / DB2 and other sources, and the data foundations that power Power BI semantic models and AI data agents. This is a hands-on engineering role in a real, non-greenfield environment with legacy ERP data that requires interpretation, cleansing and standardization.

Main Projects

•     Building a scalable, central data platform that consolidates transactional data from our AS400 (IBM i) ERP and DB2 databases.

•     Integrating additional sources — MS SQL-based applications, EDI feeds (EDI 852 / SPS Commerce), flat files and manual business files — to support analytics and reporting.

•     Designing and automating metadata-driven ETL/ELT pipelines for nightly and incremental data ingestion.

•     Implementing the medallion architecture (Bronze / Silver / Gold) with reporting-ready and AI-ready Gold datasets.

•     Implementing data quality, lineage and governance processes across the platform.

•     Preparing curated, governed data for Power BI semantic models and AI data agents (Fabric Data Agent / Copilot / custom agents).

Key Responsibilities

•     Design, develop and maintain ETL/ELT processes to extract data from AS400 DB2 and MS SQL systems.

•     Build and operate metadata-driven, parameterized pipelines supporting full and incremental loads, with retry logic, error handling and audit columns.

•     Develop transformations in Spark / PySpark and SQL to cleanse, standardize and model legacy ERP data (e.g. reconstructing product keys such as division–style–color–label).

•     Model and optimize Data Warehouse / Gold schemas (star, snowflake; fact and dimension design) for performance and scalability.

•     Create Gold views and Delta tables optimized for Power BI and governed AI access.

•     Collaborate with Business Analysts and stakeholders to translate requirements into source-to-target mappings and data solutions.

•     Implement and monitor data quality checks, alerts, pipeline-failure monitoring and remediation workflows.

•     Automate deployments and version control of data pipelines and models using Git and CI/CD across Dev / Test / Prod environments.

•     Reverse-engineer legacy ERP tables and document solutions (data dictionaries, mappings, technical notes).

•     Troubleshoot data issues and tune performance for large datasets.

Required Qualifications

•     Proven experience as a Data Engineer or similar role in a corporate environment.

•     Strong SQL skills and hands-on experience with DB2 on AS400 (IBM i) and MS SQL Server.

•     Hands-on experience with data engineering on Microsoft Fabric (or comparable Azure data engineering: Lakehouse, Delta Lake, pipelines).

•     Experience with Spark / PySpark for data transformation.

•     Solid understanding of data modeling, warehousing concepts (medallion, dimensional modeling) and performance tuning.

•     Proficiency in ETL frameworks or orchestration tools, including incremental loading and source-to-target mapping.

•     Programming experience in Python, Shell scripting or similar languages.

•     Familiarity with data governance, lineage and quality best practices.

•     Experience with Git and version control.

•     Excellent communication skills and ability to work with cross-functional teams.

•     English proficiency at C1 level or higher.

Nice-to-Have

•     Deeper Microsoft Fabric experience: OneLake, Data Pipelines / Dataflows Gen2, Fabric Notebooks, SQL Analytics Endpoint, Direct Lake.

•     Power BI awareness — semantic models, DAX, Power Query (M) — at a level enabling close collaboration with BI.

•     Tabular modeling and TMDL (model backup, Git, code review).

•     KQL / Eventhouse for monitoring and observability use cases.

•     AI-ready data modeling and awareness of Fabric Data Agent, Copilot, AI Foundry and MCP / agentic data-access concepts.

•     Governance and security in Fabric: row-level security, Microsoft Entra ID, Service Principal, least-privilege and environment separation.

•     CI/CD for data pipelines (Service Principal deployments, Azure DevOps).

•     Domain experience in apparel / fashion / supply chain: wholesale, retail reporting, sourcing, production, inventory, capacity planning, customer orders, brand / license reporting.

Who You Are

•     Comfortable working with imperfect, legacy data sources and patient enough to interpret cryptic ERP fields and reconstruct business logic.

•     Able to ask good business questions and translate technical problems into business language.

•     Thinks in terms of maintainability, scalability and data contracts — not just tables.

•     Understands the difference between raw data and curated business data, and treats data as a representation of business processes.

Hiring Process

•     Initial meeting in Polish to get acquainted and discuss your background.

•     Technical interview within two weeks, focusing on your data architecture and ETL skills.

•     Brief English conversation to assess collaboration and communication abilities.

What We Offer

•     Private medical insurance.

•     21 days of paid vacation per year (B2B contract).

•     Company-provided hardware and software.

•     Training programs and a clear career-development path.

•     Freedom to choose tools and technologies.

•     Supportive, trust-based work environment.

Tech stack

    English

    B2

    DataModeling

    advanced

    ETL/ELT

    advanced

    DB2

    advanced

    LakeHouse

    advanced

    Fabric

    advanced

    SQL

    advanced

    AS400/IBMi

    regular

    Git

    regular

    DataQuality

    regular

    Python

    regular

Office location

Check similar offers
Sorigo

Sorigo

Warszawa

Remote

Remote

Undisclosed Salary
DBT
Python
SQL
Airflow
AWS Glue
Redshift
Iceberg
MidMidB2BB2BB2B ContractB2B Contract
New
ADVERTISEMENT: Recommended by Just Join IT
Check similar offers
Sorigo

Sorigo

Warszawa

Remote

Remote

Undisclosed Salary
DBT
Python
SQL
Airflow
AWS Glue
Redshift
Iceberg
MidMidB2BB2BB2B ContractB2B Contract
New
Link Group

Link Group

Remote

Remote

35,78 - 38,43USD/h
Data
SQL
REST API
GCP
Kubernetes
Terraform
CI/CD
MidMidB2BB2BFull-timeFull-time
New
Andersen

Andersen

Remote

Remote

2 800 - 5 000USD/month
SQL
Python
DBT
MidMidAnyAnyFull-timeFull-time
New
B3 Consulting Poland

B3 Consulting Poland

Warszawa

Remote

Remote

Undisclosed Salary
dbt and layered ELT modelling
BigQuery
GCP (Cloud Run, performance tuning)
Python
GTFS static
GTFS Real-time data formats
Qlik Sense
Git / GitHub
MidMidB2BB2BFull-timeFull-time
New
Sigma Software

Sigma Software

Remote

Remote

Undisclosed Salary
Python
SQL
AWS
Data Warehousing
MidMidPermanent, B2BPermanent, B2BFull-timeFull-time
New
ADVERTISEMENT: Recommended by Just Join IT