Remote job
SysOps Engineer – Automation & Platform Operations
Job details
About this role
Role overview This role automates the day-to-day operation of server and platform environments, helping keep systems consistent, recoverable, and resilient. You will build and maintain scripts and workflows for provisioning, maintenance, monitoring responses, and disaster recovery, working with infrastructure and DevOps colleagues.
Responsibilities - Automate server setup, configuration, operating system maintenance, patching, and upgrades. - Build backup and recovery workflows, including recurring checks that confirm recovery is possible. - Automate disaster recovery drills, failover execution, environment synchronization, and reporting against recovery objectives. - Develop auto-recovery actions for common failures and connect them to monitoring and alerting events. - Maintain reusable scripts, configuration standards, and approved system images. - Document operational workflows, recovery procedures, and runbooks; coordinate infrastructure standards with DevOps.
Requirements - Experience in systems operations, infrastructure engineering, platform operations, or a similar field. - Strong administration skills for Linux and/or Windows servers. - Ability to script and automate with tools such as Bash, PowerShell, Python, or equivalents. - Practical experience with configuration management, provisioning, and infrastructure automation. - Knowledge of backup, disaster recovery, failover, and business continuity practices. - Skill in troubleshooting complex infrastructure issues, documenting solutions, and collaborating across technical teams.
Nice to have - A relevant degree or equivalent hands-on experience. - Experience with hybrid, on-premises, cloud, or multi-cloud environments. - Familiarity with monitoring and alerting integrations and server lifecycle management.