Remote job
Manager of Infrastructure and DevOps
Job details
About this role
Role overview A Manager of Infrastructure and DevOps is needed to lead the teams responsible for the CI/CD pipelines, staging, and production environments powering a high-RPS, distributed SaaS platform. The role combines hands-on technical leadership with people leadership across regions and time zones, ensuring the platform stays highly available while moving faster on releases. It reports into the engineering organization and partners broadly with product and adjacent engineering teams.
Responsibilities - Lead and grow a distributed team of DevOps and SRE engineers, providing technical guidance, coaching, and career development - Maintain staging and production environments with 99.9% and 99.99% availability targets respectively - Enforce rigorous CI/CD testing and reviews to eliminate release-related production issues - Continuously reduce lead time by accelerating pipelines, environments, and deployment strategies - Implement and evolve Infrastructure as Code, secure development environments, and disaster recovery procedures - Use a data-driven approach to justify decisions with concrete metrics - Build relationships with neighboring teams, initiate and honor cross-team agreements, and serve as the interface between the infrastructure function and the wider engineering environment - Contribute to hiring and strengthening the team across multiple regions
Requirements - At least three years leading a team in a fast-growing multinational IT organization - At least five years of DevOps and SRE experience maintaining high-RPS distributed products - Strong hands-on experience as a lead in complex CI/CD pipelines, Infrastructure as Code, and varied deployment strategies - Practical experience developing and maintaining large-scale cloud applications on AWS, Azure, or GCP - Experience implementing secure development environments - Ability to justify decisions using concrete metrics and a data-driven mindset - Comfort acting as the interface between the team and surrounding engineering groups
Nice to have - Familiarity with a modern stack built on .NET Core, MongoDB, Vue 2/3, AWS, Kafka, Elasticsearch, GitLab CI, Prometheus, Victoria Metrics, Jaeger, and Kubernetes
Benefits and work setup A globally distributed team of 200+ people across 30+ countries, with remote-friendly roles and hubs in New York, London, Lisbon, Costa Rica, Serbia, Armenia, Georgia, and Spain. The organization is described as highly innovative, embedding AI across areas of the business and encouraging employees to manage their own AI agents. It is a post-Series C company growing at roughly 130–150% year-over-year.