Cloud Operations & Observability Engineer
⚲ Gdańsk
Do uzgodnienia
Wymagania
- Grafana
- Splunk
- AWS
- DevOps
- Terraform
- AIOps
Opis stanowiska
Cloud Operations & Observability Engineer
Skywise is a leading digital services company, wholly owned by Airbus. We deliver a first-of-its-kind offer that seamlessly integrates Planning & Operations Control, Flight, Technical, and Ground Operations. By providing interoperable digital solutions for end-to-end operations, Skywise combines aircraft manufacturer expertise and digital know-how to transform operational complexity into predictable and profitable performance, ensuring resilient and more sustainable operations. Our product portfolio has recently expanded, opening up exciting opportunities to harmonize, optimize, and revolutionize how we operate our cloud infrastructure. We are
looking for a visionary engineer to help lead the charge. As a Cloud Operations & Observability Engineer, you will play a pivotal role in the future of our digital platform operations. You will lead the assessment, design, and execution of our next-generation monitoring, observability, and alerting strategy. You will partner with external experts and Skywise teams to design, build, and scale our
next-generation observability platform. Your mission will be to bridge the gap between different engineering cultures, optimize our cloud platform operations for the entire expanded portfolio, and inject cutting-edge capabilities, such as AI/MLdriven operations
(AIOps), into our ecosystem.
Your scope of activities:
● Observability & Telemetry Architecture
● Platform Harmonization: consolidation of fragmented monitoring tools into a single, unified control room
Your responsibilities
• Audit and Identify Synergies & Gaps: Assess the existing technical setups, telemetry infrastructure, monitoring tools, and operational workflows across both the legacy and newly developed or acquired solutions. Identify redundancies,
• inefficiencies, and best practices in how the different organizations currently operate their cloud platforms.
• Define the Future State: Design a unified, scalable roadmap for technical operations, focusing on robust monitoring, observability, and proactive alerting.
• Cost & Best Practice Optimization: Implement "best cost" practices to ensure tool rationalization and optimal cloud spend.
• Innovate with AI/ML: Integrate Artificial Intelligence and Machine Learning
• (AIOps) into the monitoring roadmap to enable predictive alerting, automated incident response, and anomaly detection.
• Tooling Harmonization: Drive the migration, consolidation, and optimization of the observability toolset across the organization.
• Culture & Methods: Define and evangelize modern Site Reliability Engineering (SRE) methodologies across the operational teams.
• Knowledge Transfer & Sustainability: Shadow external consultants to absorb architectural designs, eventually taking full internal ownership of the observability roadmap and platform governance.
Technical profile and skills:
• Observability & Monitoring Ecosystems: Deep, hands-on experience designing and managing enterprise-level setups using Splunk, Grafana, or Datadog.
• Cloud Infrastructure: Strong expertise in Amazon Web Services (AWS). An active AWS Certification (e.g., AWS Certified DevOps Engineer, AWS Certified Solutions Architect) is highly preferred.
• AWS Native Monitoring: Proven experience with AWS CloudWatch and related native AWS governance/monitoring tools.
• Infrastructure as Code (IaC): Familiarity with tools like Terraform or CloudFormation to ensure observability-as-code.
• Proven experience on the field of platforms operations.
• Forward-Thinking Mindset: A strong interest or practical experience in applying Machine Learning/AI to operational data (AIOps).
• Collaborative Leadership: Ability to work across cross-functional, newly merged teams to drive consensus and execute a shared vision.
• Ability to communicate needs to senior management, explaining business value
What We Offer:
• Stable, full-time employment contract
• Flexible working hours with a hybrid model (3 days per week in our Gdansk office at Olivia Business Centre)
• Training and development opportunities to support your career growth (3000 PLN gross per year after the probationary period)
• Co-funding for meals and commuting
• Access to the latest knowledge and technologies enabling professional development
• Opportunity to work on international projects and collaborate with global teams, including occasional international travel
• Private medical coverage
• Sport card
• Mental health support platform
• Life insurance
• Employee Stock Ownership Plan
• Employee referral program bonus (5000 PLN)
Recruitment Roadmap:
• HR screening (online)
• Technical interview
• Interview w/ Head of Digital Platform Operations
• Final interview w/ HR Business Partner (onsite)
Skywise is a leading digital services company, wholly owned by Airbus. We deliver a first-of-its-kind offer that seamlessly integrates Planning & Operations Control, Flight, Technical, and Ground Operations. By providing interoperable digital solutions for end-to-end operations, Skywise combines aircraft manufacturer expertise and digital know-how to transform operational complexity into predictable and profitable performance, ensuring resilient and more sustainable operations. Our product portfolio has recently expanded, opening up exciting opportunities to harmonize, optimize, and revolutionize how we operate our cloud infrastructure. We are
looking for a visionary engineer to help lead the charge. As a Cloud Operations & Observability Engineer, you will play a pivotal role in the future of our digital platform operations. You will lead the assessment, design, and execution of our next-generation monitoring, observability, and alerting strategy. You will partner with external experts and Skywise teams to design, build, and scale our
next-generation observability platform. Your mission will be to bridge the gap between different engineering cultures, optimize our cloud platform operations for the entire expanded portfolio, and inject cutting-edge capabilities, such as AI/MLdriven operations
(AIOps), into our ecosystem.
Your scope of activities:
● Observability & Telemetry Architecture
● Platform Harmonization: consolidation of fragmented monitoring tools into a single, unified control room
Your responsibilities
• Audit and Identify Synergies & Gaps: Assess the existing technical setups, telemetry infrastructure, monitoring tools, and operational workflows across both the legacy and newly developed or acquired solutions. Identify redundancies,
• inefficiencies, and best practices in how the different organizations currently operate their cloud platforms.
• Define the Future State: Design a unified, scalable roadmap for technical operations, focusing on robust monitoring, observability, and proactive alerting.
• Cost & Best Practice Optimization: Implement "best cost" practices to ensure tool rationalization and optimal cloud spend.
• Innovate with AI/ML: Integrate Artificial Intelligence and Machine Learning
• (AIOps) into the monitoring roadmap to enable predictive alerting, automated incident response, and anomaly detection.
• Tooling Harmonization: Drive the migration, consolidation, and optimization of the observability toolset across the organization.
• Culture & Methods: Define and evangelize modern Site Reliability Engineering (SRE) methodologies across the operational teams.
• Knowledge Transfer & Sustainability: Shadow external consultants to absorb architectural designs, eventually taking full internal ownership of the observability roadmap and platform governance.
Technical profile and skills:
• Observability & Monitoring Ecosystems: Deep, hands-on experience designing and managing enterprise-level setups using Splunk, Grafana, or Datadog.
• Cloud Infrastructure: Strong expertise in Amazon Web Services (AWS). An active AWS Certification (e.g., AWS Certified DevOps Engineer, AWS Certified Solutions Architect) is highly preferred.
• AWS Native Monitoring: Proven experience with AWS CloudWatch and related native AWS governance/monitoring tools.
• Infrastructure as Code (IaC): Familiarity with tools like Terraform or CloudFormation to ensure observability-as-code.
• Proven experience on the field of platforms operations.
• Forward-Thinking Mindset: A strong interest or practical experience in applying Machine Learning/AI to operational data (AIOps).
• Collaborative Leadership: Ability to work across cross-functional, newly merged teams to drive consensus and execute a shared vision.
• Ability to communicate needs to senior management, explaining business value
What We Offer:
• Stable, full-time employment contract
• Flexible working hours with a hybrid model (3 days per week in our Gdansk office at Olivia Business Centre)
• Training and development opportunities to support your career growth (3000 PLN gross per year after the probationary period)
• Co-funding for meals and commuting
• Access to the latest knowledge and technologies enabling professional development
• Opportunity to work on international projects and collaborate with global teams, including occasional international travel
• Private medical coverage
• Sport card
• Mental health support platform
• Life insurance
• Employee Stock Ownership Plan
• Employee referral program bonus (5000 PLN)
Recruitment Roadmap:
• HR screening (online)
• Technical interview
• Interview w/ Head of Digital Platform Operations
• Final interview w/ HR Business Partner (onsite)
🔍 Dekoder Ogłoszenia
🔴
visionary engineer to help lead the charge
Szukamy kogoś, kto przejmie inicjatywę i będzie wyznaczał kierunki, ale bez gwarancji formalnego przywództwa.
🔴
bridge the gap between different engineering cultures
Może oznaczać konieczność rozwiązywania konfliktów lub radzenia sobie z nieefektywną komunikacją między zespołami.
🔴
optimize our cloud platform operations for the entire expanded portfolio
Praca nad istniejącą, potencjalnie złożoną i nieoptymalną infrastrukturą, która właśnie się powiększyła.
🔴
inject cutting-edge capabilities, such as AI/ML-driven operations (AIOps), into our ecosystem
Wdrożenie nowych, zaawansowanych technologii może oznaczać niepewność, brak gotowych rozwiązań i konieczność eksperymentowania.
🔴
Platform Harmonization: consolidation of frag
Fragment ogłoszenia jest niekompletny, co sugeruje pośpiech w jego tworzeniu lub brak dopracowania.