Bachelor’s degree in Computer Science, Engineering, Mathematics or a related discipline, or equivalent professional experience
Experience developing, administering and supporting Nagios or similar enterprise monitoring platforms
Experience with Prometheus, Grafana & other monitoring tools beneficial
Experience designing, building and maintaining CI/CD pipelines
Experience deploying and managing applications in Kubernetes using Helm
Strong automation and configuration management experience using Ansible
Scripting or software development experience, preferably with Python
Experience working with Git-based development and deployment workflows
Experience with Linux systems administration in a medium to large enterprise environment
An understanding of infrastructure as code principles and experience migrating manually managed infrastructure and services towards an automated, version-controlled model
Strong troubleshooting skills and the ability to support production systems in a reliable and repeatable manner
Good communication skills and the ability to work effectively with globalised engineering and operational teams
What you'll be doing
You will be responsible for developing, deploying and supporting systems that monitor and manage network devices, servers and applications
You will also work on projects to modernise our deployment processes
Improve our monitoring and alerting capabilities
Migrate existing service towards an infrastructure as code model