The role that notices the disk filling up before a customer does. You will look after monitoring, backups and access control across the fleet, and be the person who actually tests that a restore works.
What you will do
- Run monitoring and alerting, and keep the noise low enough that alerts still mean something
- Own the backup schedule and prove it by restoring from it regularly
- Manage accounts, SSH keys and least-privilege access across the estate
- Handle escalations from support that need root
- Keep runbooks current so the next person does not have to guess
What we are looking for
- Day-to-day Linux administration in production
- Practical backup and restore experience, not just backup experience
- Shell scripting, and enough of a language like Python to automate the repetitive parts
- Careful with production access and able to explain what you changed and why
Nice to have
- cPanel/WHM at scale
- Prometheus/Grafana or Zabbix
- DNS and mail deliverability troubleshooting
Apply for this role
No positions are open at the moment. The roles listed here are the kinds of work we hire for, published so you can see what we look for. Applications are closed and this form cannot be submitted yet — please check back, or get in touch if you would like us to keep you in mind.
