01 Zakres zadań
- Design and implement metrics, logging, tracing and telemetry for cloud and edge components.
- Build monitoring pipelines, dashboards and alerting frameworks for deployment health, device status and store operations.
- Define and maintain SLIs, SLOs and reliability metrics for critical deployment services.
- Create operational analytics for rollout performance, failure detection and capacity planning.
- Integrate telemetry from cloud services, edge devices, deployment agents and infrastructure into centralized monitoring platforms.
- Develop automated detection, root-cause analysis, self-service dashboards and reporting for engineering, support and operations teams.
- Drive observability standards and monitoring strategies across the Retail Platform.
- Work in a hybrid model near Kraków: 60% of the week, or three days, at offices or local sites and the remaining time remotely.
