Remote job
Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms
Job details
About this role
Role overview A Site Reliability Engineering role open to candidates from intermediate through senior staff levels, with a focus on infrastructure platforms. The position is remote across Canada, the United Kingdom, and the United States, and involves operating and evolving the foundational systems that production services depend on.
Responsibilities - Operate and improve infrastructure platforms supporting production workloads - Build automation for deployment, scaling, and recovery of platform services - Participate in incident response, on-call rotations, and post-incident reviews - Contribute to observability, capacity planning, and reliability goals - Collaborate with product engineering teams on platform capabilities and roadmaps
Requirements - Production experience as an SRE or infrastructure engineer at a level appropriate to the applied-for band - Proficiency with operating systems, networking, and at least one major cloud provider - Familiarity with infrastructure-as-code and configuration management practices - Comfort with on-call rotations and production troubleshooting under pressure - Strong written communication for a distributed, asynchronous environment