Data Engineer
⚲ Warszawa
13 500 - 17 500 PLN netto (B2B)
Wymagania
- Data Engineering
- ELT pipeline
- Data Quality
- data observability
Opis stanowiska
Ready to reshape a data platform from the ground up? Join Us!
As a leading AWS Partner in Europe, we support organizations at every stage of their cloud journey. Our strength comes from deep experience in cloud-native architectures. By joining our team, you'll have the opportunity to design and build scalable data platforms on AWS, work hands-on with modern data stack technologies, and directly shape how businesses in regulated industries process, store, and make sense of their data.
Join our Data & AI team and work on a project where the decisions you make today become the foundation others build on tomorrow. As a Data Engineer, you'll join an insurance industry data platform project - an AWS-native solution that recently shipped its MVP and is now entering a phase of serious architectural evolution. You'll be rebuilding it the right way: from the ground up, with the right tools, the right patterns, and full ownership over what you ship.
What we're building right now:• Migrating from direct-to-Redshift streaming to a proper data lake architecture
• Introducing batch loading alongside existing streaming pipelines
• Replacing Step Functions with Airflow (AWS MWAA) as the orchestration layer
• Moving workloads from ECS to EKS for better scalability and control
• Working with Kafka + Debezium for CDC, dltHub for ingestion, dbt + Cosmos for transformations, and OpenMetadata for data governance
We work in a "we build it, we run it" model - every engineer on the team can deploy their solutions to production without handing off to anyone else.
Your Responsibilities:
• Design and implement the new data lake architecture on AWS
• Build and maintain ELT pipelines using dbt, dltHub, and Airflow
• Migrate existing orchestration from Step Functions to Airflow (MWAA)
• Ensure data quality, observability, and lineage across the platform
• Deploy your own solutions to production - no handoffs, full ownership
• Collaborate closely with a small, senior engineering team and contribute to technical decisions
Requirements:
• 2+ years of experience in data engineering, with production-grade delivery
• Hands-on experience with dbt and at least one modern orchestration tool (Airflow preferred)
• Solid Python skills and experience building and operating data pipelines on AWS
• Familiarity with data lake concepts and batch/streaming architectures
• Comfort with Infrastructure as Code (Terraform or similar)
• Self-driven and able to operate with high autonomy - you don't wait to be told what to do next
• Fluency in both English (C1) and Polish (C1), written and spoken
Nice to have:
• Experience with Kafka, Debezium, or other CDC tooling
• Familiarity with dltHub or similar ingestion frameworks
• Hands-on experience with Amazon Redshift and data warehouse optimization
• Background with OpenMetadata or other data catalog/governance tools
• Experience with healthcare datasets (FHIR, OMOP, DICOM)
Benefits:
• Continuous learning and growth – we invest in your development through internal training, knowledge-sharing sessions, and hands-on learning from real projects. We also support growth beyond the organization by actively engaging in the AWS Community, including conference talks and industry events.
• Training budget – dedicated funds to support certifications, courses, and professional development aligned with your career goals.
• Medical care package – comprehensive private healthcare to help you stay healthy and focused.
• Multisport card co-financing – because staying active matters, both in and outside of work.
• Language learning platform access – improve your language skills at your own pace, whenever it suits you.
• Flexible working hours & remote work – work when and where you’re most productive.
• Company events – time for integration and some fun.
Stay up to date with us: Follow our blog regularly and participate in the events we organize.
As a leading AWS Partner in Europe, we support organizations at every stage of their cloud journey. Our strength comes from deep experience in cloud-native architectures. By joining our team, you'll have the opportunity to design and build scalable data platforms on AWS, work hands-on with modern data stack technologies, and directly shape how businesses in regulated industries process, store, and make sense of their data.
Join our Data & AI team and work on a project where the decisions you make today become the foundation others build on tomorrow. As a Data Engineer, you'll join an insurance industry data platform project - an AWS-native solution that recently shipped its MVP and is now entering a phase of serious architectural evolution. You'll be rebuilding it the right way: from the ground up, with the right tools, the right patterns, and full ownership over what you ship.
What we're building right now:• Migrating from direct-to-Redshift streaming to a proper data lake architecture
• Introducing batch loading alongside existing streaming pipelines
• Replacing Step Functions with Airflow (AWS MWAA) as the orchestration layer
• Moving workloads from ECS to EKS for better scalability and control
• Working with Kafka + Debezium for CDC, dltHub for ingestion, dbt + Cosmos for transformations, and OpenMetadata for data governance
We work in a "we build it, we run it" model - every engineer on the team can deploy their solutions to production without handing off to anyone else.
Your Responsibilities:
• Design and implement the new data lake architecture on AWS
• Build and maintain ELT pipelines using dbt, dltHub, and Airflow
• Migrate existing orchestration from Step Functions to Airflow (MWAA)
• Ensure data quality, observability, and lineage across the platform
• Deploy your own solutions to production - no handoffs, full ownership
• Collaborate closely with a small, senior engineering team and contribute to technical decisions
Requirements:
• 2+ years of experience in data engineering, with production-grade delivery
• Hands-on experience with dbt and at least one modern orchestration tool (Airflow preferred)
• Solid Python skills and experience building and operating data pipelines on AWS
• Familiarity with data lake concepts and batch/streaming architectures
• Comfort with Infrastructure as Code (Terraform or similar)
• Self-driven and able to operate with high autonomy - you don't wait to be told what to do next
• Fluency in both English (C1) and Polish (C1), written and spoken
Nice to have:
• Experience with Kafka, Debezium, or other CDC tooling
• Familiarity with dltHub or similar ingestion frameworks
• Hands-on experience with Amazon Redshift and data warehouse optimization
• Background with OpenMetadata or other data catalog/governance tools
• Experience with healthcare datasets (FHIR, OMOP, DICOM)
Benefits:
• Continuous learning and growth – we invest in your development through internal training, knowledge-sharing sessions, and hands-on learning from real projects. We also support growth beyond the organization by actively engaging in the AWS Community, including conference talks and industry events.
• Training budget – dedicated funds to support certifications, courses, and professional development aligned with your career goals.
• Medical care package – comprehensive private healthcare to help you stay healthy and focused.
• Multisport card co-financing – because staying active matters, both in and outside of work.
• Language learning platform access – improve your language skills at your own pace, whenever it suits you.
• Flexible working hours & remote work – work when and where you’re most productive.
• Company events – time for integration and some fun.
Stay up to date with us: Follow our blog regularly and participate in the events we organize.
🔍 Dekoder Ogłoszenia
🔴
Ready to reshape a data platform from the ground up?
Oczekuje się, że będziesz aktywnie uczestniczyć w znaczących zmianach i przebudowie istniejącej platformy, a nie tylko utrzymaniu jej.
🔴
you'll be rebuilding it the right way: from the ground up, with the right tools, the right patterns, and full ownership over what you ship.
Może to oznaczać, że obecne rozwiązanie jest dalekie od ideału i wymaga gruntownej przebudowy, a Ty będziesz odpowiedzialny za ten proces.
🔴
the decisions you make today become the foundation others build on tomorrow.
Twoje decyzje będą miały długoterminowy wpływ na architekturę, co wiąże się z dużą odpowiedzialnością i potencjalnym naciskiem na podejmowanie 'właściwych' decyzji.
🔴
we work in a "we build it, we run it" model - every engineer on the team can deploy
Oczekuje się, że będziesz nie tylko tworzyć, ale także wdrażać i monitorować swoje rozwiązania, co może oznaczać szerszy zakres obowiązków niż typowy Data Engineer.