← Back to jobs

Remote job

Sr Site Reliability Engineer

DevOps Full-time Permanent US and Canada

Job details

Not specified Salary
US and Canada Eligibility
Senior Experience
Full-time Employment

About this role

Role overview A senior engineering role focused on cloud platform reliability, joining a Platform Infrastructure team responsible for a microservices-based software solution built on modern orchestration technologies. The position emphasizes leading the engineering organization toward standardized, automated infrastructure and service provisioning, while operating within a culture that values flexibility, trust, and continuous learning. The work blends hands-on engineering with cross-team technical leadership and on-call incident response.

Responsibilities - Architect long-term technical solutions and cross-team mechanisms that move the platform toward measurable reliability goals. - Define and drive a roadmap for self-service, automated, scalable, and observable infrastructure services delivered as a product. - Align stakeholders and lead execution of the Platform Infrastructure team's roadmap alongside senior engineers across the organization. - Provide expert guidance and review during engineering design sessions for teams onboarding to the shared platform. - Aggressively reduce manual toil through automation across provisioning, deployment, and operations. - Build and maintain monitoring and alerting for the platform, and participate in an on-call rotation for incident response.

Requirements - Substantial experience designing technical solutions for reliability at scale in a cloud-native, microservices environment. - Demonstrated ability to lead roadmap-level initiatives and influence engineering practices across multiple teams. - Hands-on expertise with AWS services such as S3, EC2, and RDS, including EKS-managed Kubernetes clusters. - Practical knowledge of service mesh technologies (Istio) and Infrastructure as Code tools such as Terraform or AWS CDK. - Familiarity with continuous integration in GitLab and continuous delivery pipelines using tooling like ArgoCD. - Experience with observability platforms such as Datadog for monitoring and alerting.

Benefits and work setup - Culture built around flexibility, trust, and continual learning, with explicit emphasis on diversity and inclusion as guiding values. - On-call rotation is part of the role, suggesting structured incident response and operational ownership expectations. - Position framed as a leadership-track opportunity within the platform infrastructure function.

Skills detected in the listing

Information SecurityAWSKubernetesTerraform
Detected Oct 8, 2026
Last verified Oct 8, 2026

Hidden Jobs Access

Unlock application links

Read the full job details for free. An active Hidden Jobs Access subscription is required to open the original application link.

Weekly

FREE $6.99/week after trial
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
  • Cancel anytime before day 7

Monthly

$35.99 $17.99 /month
  • 35% cheaper than weekly
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching

Lifetime

$99.99 $49.99 /forever
  • One-time payment
  • Original application links
  • Daily or weekly job alerts
  • Premium filters and CV matching
Hidden Jobs gives subscribers direct access to original application links
Offer ends in 00:00:00 Your profile-fit rate expires at midnight