Remote job
Sr. Technical Program Manager
Job details
About this role
Role overview
A senior Technical Program Manager is needed to lead Site Reliability Engineering (SRE) initiatives that strengthen infrastructure reliability, scalability, and operational efficiency across a large SaaS platform. This individual contributor role partners with engineering, product, and operations leaders to drive SRE best practices end to end. The position reports to the Senior Director of Cloud and Production Engineering and is fully remote, though the role is not eligible for hire in several U.S. states.
Responsibilities
- Own end-to-end delivery of SRE programs spanning incident management, release management, observability, automation, and capacity planning. - Coordinate across software engineering, SRE, and DevOps teams to align priorities, unblock work, and track technical deliverables, timelines, and cross-team dependencies. - Build dashboards, status reports, and executive updates that surface program milestones, risks, and outcomes in business-friendly language. - Partner with Incident Commanders on post-incident reviews, driving root cause analysis and ensuring preventive actions are implemented to reduce downtime. - Translate technical detail into business impact for stakeholders, leadership, and cross-functional partners.
Requirements
- 8+ years of professional experience in high-tech, including 5+ years in technical program management or SRE supporting SaaS platform and infrastructure. - Track record of managing technically complex programs involving 25+ team members and delivering under tight deadlines. - Hands-on familiarity with SRE principles (observability, incident response, infrastructure automation), distributed systems, and major cloud platforms such as AWS, Azure, or GCP. - Experience with Kubernetes, CI/CD pipelines, infrastructure-as-code tools like Terraform or Ansible, and monitoring stacks such as Prometheus and Grafana. - Proficiency with project management tools (Jira, Asana, Microsoft Project) and agile or scrum delivery frameworks. - Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
Nice to have
- Experience defining and rolling out SLOs and SLIs for large-scale systems. - Comfort setting OKRs and managing multiple parallel deliverables. - Background working in cloud or production engineering organizations.