Remote job
Senior Site Reliability Engineer (SRE) – Infrastructure & Systems (Remote)
Job details
About this role
Role overview This Senior Site Reliability Engineer role owns the production infrastructure of an enterprise-grade proxy platform that delivers tens of millions of IPs worldwide and targets 99.97% network uptime. The engineer will lead the on-premises migration from Docker Swarm to Kubernetes, maintain high availability across hundreds of servers and roughly fifty services, and partner closely with backend engineers.
Responsibilities - Own and evolve the production platform, including hundreds of servers and approximately fifty services. - Lead the migration from Docker Swarm to a self-hosted Kubernetes setup on bare metal. - Drive observability improvements in cooperation with the development team and enforce infrastructure-as-code and CI/CD reliability practices. - Participate in an on-call rotation alongside backend engineers and lead incident response, including post-mortems and systematic remediation. - Build platform tooling that reduces infrastructure toil and improves developer experience.
Requirements - Hands-on experience operating highly available infrastructure at real production scale, spanning hundreds of servers and dozens of services. - Practical Kubernetes experience in self-hosted or bare-metal environments. - Confidence with infrastructure-as-code tooling and practices. - Solid scripting and development skills. - A proactive mindset that surfaces problems early and keeps the team informed without prompting.
Nice to have - Prior experience leading at least one major infrastructure migration from planning through stabilization. - Familiarity with Python or Go, since backend services are written in Python and edge services in Go. - Exposure to proxy or other networking-heavy infrastructure. - Background in small teams where developers share infrastructure responsibility. - Familiarity with edge clusters or split compute and edge architectures.
Benefits and work setup - Gross salary starting at 32,000 PLN per month plus a quarterly KPI-based performance bonus, with openness to discuss different levels based on skills and experience. - Remote work setup, collaborating with a strong backend team that shares infrastructure ownership. - Private health insurance, gym allowance, and a wellness app. - Forty-plus internal learning options, plus support for external conferences, mentorship, and year-round knowledge sharing. - Team events and an overseas workation to mark milestones together.