GCP Data Platform Engineer - Automation & Innovation Department
⚲ Warszawa, Mokotów
Do uzgodnienia
Wymagania
- BigQuery
- Dataflow
- Dataproc
- Composer
- Python
- Kafka
- Terraform
- Docker
- Kubernetes
- GitLab CI/CD
- Looker
- Vertex AI
- Spark
- Scala
- Java
Opis stanowiska
Nasze wymagania:
3+ years of experience as a Data Platform Engineer in a data-driven environment.
Experience in developing enterprise-ready solutions based on GCP data services (BigQuery, Cloud Storage, Pub/Sub, Dataproc, Composer, Cloud Run, Looker, Vertex AI).
Experience in large-scale data migration or cloud transformation projects.
Experience with modern data platform patterns, including data lakehouse architectures on GCP (Cloud Storage + BigQuery).
Hands-on experience with Infrastructure-as-Code (IaC) tools, including Terraform/Terragrunt.
Proficiency in Python.
Experience with Linux, Docker/Kubernetes, and GitLab CI/CD pipelines.
Very good command of English (spoken and written).
Strong communication skills with the ability to explain complex technical concepts to business stakeholders.
Mile widziane:
Degree in Computer Science.
Experience with Spark.
Scala or Java knowledge.
Knowledge of data governance, metadata, and data quality tools.
Experience collaborating with business stakeholders.
O projekcie:
We’re moving analytics from on-premise to GCP and building our data architecture and data model from the ground up.
We integrate diverse, high-volume data sources, design streaming and batch processing layers, implement data governance, lineage, data quality, and data security, and set up CI/CD and monitoring/SLOs.
Strong focus on business value creation and CX of our customers.
Zakres obowiązków:
Develop reusable frameworks for data processing and testing on GCP.
Build and maintain batch and streaming data ingestion pipelines from various sources into GCP.
Implement automated tests and data quality checks for data pipelines.
Collaborate with analysts and data scientists to deliver reliable, well-documented datasets.
Monitor, optimize, and secure data pipelines in line with data governance and compliance standards.
Oferujemy:
Join a new, strategic data transformation project, where analytics are being moved from on-premise to GCP.
Build data architecture and data models from the ground up.
Strong focus on business value creation and customer experience.
Actively shape the standards, patterns, and long-term direction of the data platform.
3+ years of experience as a Data Platform Engineer in a data-driven environment.
Experience in developing enterprise-ready solutions based on GCP data services (BigQuery, Cloud Storage, Pub/Sub, Dataproc, Composer, Cloud Run, Looker, Vertex AI).
Experience in large-scale data migration or cloud transformation projects.
Experience with modern data platform patterns, including data lakehouse architectures on GCP (Cloud Storage + BigQuery).
Hands-on experience with Infrastructure-as-Code (IaC) tools, including Terraform/Terragrunt.
Proficiency in Python.
Experience with Linux, Docker/Kubernetes, and GitLab CI/CD pipelines.
Very good command of English (spoken and written).
Strong communication skills with the ability to explain complex technical concepts to business stakeholders.
Mile widziane:
Degree in Computer Science.
Experience with Spark.
Scala or Java knowledge.
Knowledge of data governance, metadata, and data quality tools.
Experience collaborating with business stakeholders.
O projekcie:
We’re moving analytics from on-premise to GCP and building our data architecture and data model from the ground up.
We integrate diverse, high-volume data sources, design streaming and batch processing layers, implement data governance, lineage, data quality, and data security, and set up CI/CD and monitoring/SLOs.
Strong focus on business value creation and CX of our customers.
Zakres obowiązków:
Develop reusable frameworks for data processing and testing on GCP.
Build and maintain batch and streaming data ingestion pipelines from various sources into GCP.
Implement automated tests and data quality checks for data pipelines.
Collaborate with analysts and data scientists to deliver reliable, well-documented datasets.
Monitor, optimize, and secure data pipelines in line with data governance and compliance standards.
Oferujemy:
Join a new, strategic data transformation project, where analytics are being moved from on-premise to GCP.
Build data architecture and data models from the ground up.
Strong focus on business value creation and customer experience.
Actively shape the standards, patterns, and long-term direction of the data platform.
🔍 Dekoder Ogłoszenia
🔴
Automation & Innovation Department
Może oznaczać, że dział jest w fazie rozwoju, a procesy i narzędzia nie są jeszcze w pełni ustabilizowane.
🔴
enterprise-ready solutions
Wymaga tworzenia rozwiązań, które są skalowalne, bezpieczne i spełniają standardy korporacyjne, co może oznaczać dużą odpowiedzialność i rygorystyczne wymagania.
🔴
building our data architecture and data model from the ground up
Oznacza, że projekt jest na wczesnym etapie, co może wiązać się z niepewnością co do ostatecznego kształtu architektury i koniecznością częstych zmian.
🔴
Strong focus on business value creation and CX of our customers
Może sugerować, że oczekuje się od inżyniera nie tylko technicznych umiejętności, ale także zrozumienia potrzeb biznesowych i wpływu pracy na użytkownika końcowego, co może oznaczać dodatkowe wymagania poza czysto technicznymi.
🟡
Develop reusable frameworks for data processing and testing on GCP
Choć brzmi pozytywnie, może oznaczać, że duża część pracy będzie polegać na tworzeniu narzędzi i szablonów, a nie bezpośrednio na analizie danych czy budowaniu konkretnych rozwiązań analitycznych.