← Back to jobs

Remote job

Senior Site Reliability Engineer

DevOps Full-time Permanent Canada

Job details

$151,000/year Salary
Canada Eligibility
Senior Experience
Full-time Employment

About this role

Role overview Join a Platform SRE team to build and operate the infrastructure, tools, and paved roads that help developers deliver scalable, secure, and reliable software. The role spans infrastructure automation, observability, developer enablement, and system reliability, with a focus on reducing toil and creating self-service capabilities across the engineering organization.

Responsibilities - Design, build, and scale production environments using AWS and Terraform, driving architectural decisions that improve long-term maintainability - Lead efforts to improve platform resilience through failure-based testing, automated recovery strategies, and proactive capacity planning - Own the design and delivery of reusable platform components and self-service tools that streamline the developer experience - Define and evolve observability standards including system metrics, distributed tracing, and SLO frameworks - Drive projects end to end from scoping and estimation through planning, execution, and rollout - Mentor engineers, uphold infrastructure quality, and shape best practices and standards used across the organization - Engage in technical design discussions, provide guidance, and adapt strategies based on team input - Participate in a low-volume on-call rotation

Requirements - 6+ years of experience in SRE, DevOps, Cloud Engineering, or Software Development roles - Hands-on experience operating production environments in AWS - Proficiency in Go or Python with experience building production-grade automation, tooling, or libraries - Strong experience with Terraform for infrastructure as code - Experience with container orchestration platforms such as ECS or Kubernetes - Familiarity with CI/CD tools such as GitHub Actions - Solid understanding of observability practices including system metrics, distributed tracing, and SLOs - Experience with failure-based testing approaches and automated recovery strategies - Strong written and verbal communication and leadership skills

Nice to have - Experience with microservices architectures - Exposure to Kafka or other event streaming systems - Background building internal developer platforms or self-service infrastructure - Familiarity with systems security, compliance requirements, or hardening practices

Benefits and work setup - 100% medical, dental, and vision coverage - Flexible paid time off - Annual home office stipend and WeWork access - Mental and physical health wellness programs - Remote-first, highly inclusive culture with opportunities for advancement - Competitive compensation, with U.S. base salary for this position ranging from $151,000 to $201,000 depending on location, and adjusted ranges for other Canadian markets

Skills detected in the listing

PythonGoInformation SecurityAWSKubernetesTerraform
Detected Oct 9, 2026
Last verified Oct 9, 2026

Hidden Jobs Access

Unlock application links

Read the full job details for free. An active Hidden Jobs Access subscription is required to open the original application link.

Weekly

FREE $6.99/week after trial
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
  • Cancel anytime before day 7

Monthly

$35.99 $17.99 /month
  • 35% cheaper than weekly
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching

Lifetime

$99.99 $49.99 /forever
  • One-time payment
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
Hidden Jobs gives subscribers direct access to original application links
Offer ends in 00:00:00 Your profile-fit rate expires at midnight