01 Zakres zadań
You will
- Monitor and support critical production applications, responding to incidents and service requests within agreed SLAs
- Troubleshoot production issues using logs, monitoring dashboards, SQL data and application behaviour
- Investigate incidents in Unix/Linux environments using standard command-line tools.
- Coordinate recovery activities with development, engineering and infrastructure teams.
- Participate in root cause analysis and post-incident reviews
- Improve monitoring, reduce alert noise and increase operational resilience
- Create and maintain operational documentation, runbooks and knowledge articles
- Automate repetitive operational tasks using Bash, PowerShell or Python
- Develop simple tools and scripts that improve team efficiency
- Use AI productivity tools such as Copilot or Claude to support automation, documentation and operational analysis
- Support deployments, patching, disaster recovery exercises and platform maintenance
- Review operational readiness of changes and identify potential production support risks
- Communicate service status and recovery progress to relevant stakeholders
